Oliver Habryka

This stream works on infrastructure for AI safety research: AI tools that give safety researchers uplift, mechanism design and product development for funder coordination, and AI policy scenarios and proposals.

Stream overview

  1. Work with people on providing uplift to safety researchers by developing better AI tools, and integrating AI into their research and thinking flow
  2. Work with people on mechanism design and product development for funder coordination (https://docs.google.com/document/d/1GUs4_wSaaU8y8IWo5qMSmZoc6mDgpgD45fqssrlDHvc/edit?tab=t.0)
  3. More AI policy scenarios and proposals similar to AI 2040

Mentors

Oliver Habryka
Lightcone Infrastructure
,
CEO
SF Bay Area
No items found.

Oliver Habryka is the chief executive officer of Lightcone Infrastructure, the nonprofit behind the LessWrong forum and the Lighthaven venue. Habryka also leads Lightcone Commons, the organization's platform for funding ambitious philanthropic projects.

Read more

Mentorship style

Fellows we are looking for

Project selection

Streams

The Winter 2027 cohort offers a wide range of research streams led by experts across AI alignment, interpretability, governance, and safety. Each stream provides its own research agenda, methodology, and mentorship focus.

London
Empirical
Interpretability
London
Interpretability, Red-Teaming, Monitoring
London
Monitoring, Adversarial Robustness, Control, Model Organisms, Red-Teaming, Dangerous Capability Evals, Safeguards
New York City
Dangerous Capability Evals, Control, Strategy & Forecasting, Policy & Governance, Scalable Oversight, Agent Foundations
SF Bay Area
Empirical
Theory
Dangerous Capability Evals, Adversarial Robustness, Security, Red-Teaming, Scalable Oversight
London
Control, Scheming & Deception, Dangerous Capability Evals, Monitoring
Washington, D.C.
Policy & Governance, Strategy & Forecasting
Oxford
AI Welfare
SF Bay Area
Control, Model Organisms, Scheming & Deception, Strategy & Forecasting
SF Bay Area
Interpretability
Tübingen
Dangerous Capability Evals, Agent Foundations, Adversarial Robustness, Monitoring, Scalable Oversight, Scheming & Deception
SF Bay Area
Dangerous Capability Evals, Policy & Governance
New York City
Monitoring, Dangerous Capability Evals, Scalable Oversight, Safeguards
SF Bay Area
Strategy & Forecasting, Policy & Governance
Montreal
Agent Foundations, Dangerous Capability Evals, Monitoring, Control, Red-Teaming, Scalable Oversight
SF Bay Area
Control, Model Organisms, Red-Teaming, Scheming & Deception