Owain has a broad interest in AI alignment and reducing AGI risk. He is investigating dangerous capabilities and the emergence of misalignment in LLMs, along with self-awareness and latent reasoning. Owain previously worked on AI deception (How to Catch an AI Liar), truthfulness (TruthfulQA), and the Reversal Curse. Owain runs an independent AI Safety non-profit, based at Constellation in Berkeley. He previously worked at the University of Oxford and at Ought. He has mentored 30+ junior AI Safety researchers through MATS and other programs.
Jan worked as a software developer for over a decade before shifting to AI safety in 2023. He is an ARENA and Astra Fellowship alumni, interested in anything related to out-of-context reasoning in LLMs.
The Winter 2026 cohort offers a wide range of research streams led by experts across AI alignment, interpretability, governance, and safety. Each stream provides its own research agenda, methodology, and mentorship focus.