Micah Carroll

OpenAI

Member of Technical Staff, Safety Systems

Links

Focus

AI Control and Monitoring, Capability and Propensity Evaluations, Misalignment Science, Alignment Training Methods, Structural Risk and Societal Dynamics

Micah is a researcher on OpenAI’s safety team interested in AI deception, scalable oversight, and monitorability. He is on leave from a UC Berkeley PhD focused on AI alignment with influenceable humans, AI manipulation from RL training, and recommender-system effects.