Oliver Habryka

This stream works on infrastructure for AI safety research: AI tools that give safety researchers uplift, mechanism design and product development for funder coordination, and AI policy scenarios and proposals.

Stream overview

  1. Work with people on providing uplift to safety researchers by developing better AI tools, and integrating AI into their research and thinking flow
  2. Work with people on mechanism design and product development for funder coordination (https://docs.google.com/document/d/1GUs4_wSaaU8y8IWo5qMSmZoc6mDgpgD45fqssrlDHvc/edit?tab=t.0)
  3. More AI policy scenarios and proposals similar to AI 2040

Mentors

Oliver Habryka
Lightcone Infrastructure
,
CEO
SF Bay Area
No items found.

Oliver Habryka is the chief executive officer of Lightcone Infrastructure, the nonprofit behind the LessWrong forum and the Lighthaven venue. Habryka also leads Lightcone Commons, the organization's platform for funding ambitious philanthropic projects.

Read more

Mentorship style

Fellows we are looking for

Project selection

Streams

The Winter 2027 cohort offers a wide range of research streams led by experts across AI alignment, interpretability, governance, and safety. Each stream provides its own research agenda, methodology, and mentorship focus.

Grand Rapids
Agent Foundations
Washington, D.C.
Compute Infrastructure, Policy & Governance, Security
London
Dangerous Capability Evals, Compute Infrastructure, Policy & Governance, Strategy & Forecasting
Washington, D.C.
Compute Infrastructure, Security
London
Control, Monitoring
London
Control, Scheming & Deception, Dangerous Capability Evals, Model Organisms, Monitoring
SF Bay Area
Security, Dangerous Capability Evals
SF Bay Area
Dangerous Capability Evals
Boston
Adversarial Robustness, Policy & Governance, Red-Teaming, Safeguards
New York City
Control, Scalable Oversight, Red-Teaming, Model Organisms, Monitoring
SF Bay Area
Policy & Governance
SF Bay Area
Control, Monitoring, Dangerous Capability Evals
SF Bay Area
Security, Compute Infrastructure
London
Interpretability
London
Scheming & Deception, Dangerous Capability Evals, Control, Red-Teaming
SF Bay Area
Dangerous Capability Evals, Red-Teaming, Model Organisms, Control, Monitoring
Toronto
Interpretability
London
Control, Monitoring, Safeguards, Dangerous Capability Evals, Scheming & Deception
Chicago
Biorisk, Security, Safeguards
SF Bay Area
Interpretability, Agent Foundations