Micah Carroll

OpenAI

—

Member of Technical Staff, Safety Systems

Links

Focus

AI Control and Monitoring, Capability and Propensity Evaluations, Misalignment Science, Alignment Training Methods, Structural Risk and Societal Dynamics

Micah is a researcher on OpenAI’s safety team interested in AI deception, scalable oversight, and monitorability. He is on leave from a UC Berkeley PhD focused on AI alignment with influenceable humans, AI manipulation from RL training, and recommender-system effects.