← The warnings
Extinction riskFrontier labLead, Alignment Science team, Anthropic2 entries on the record
Follow
- X@janleike
- Substackaligned.substack.com
- Websitejan.leike.name
- Wikipediaen.wikipedia.org
Jan Leike
Alignment researcher who co-led OpenAI’s superalignment team, resigned in May 2024 saying safety had taken a backseat to products, and now works at Anthropic, where he maintains that supervising superhuman models remains unsolved.
Roles
- 2021 – May 2024Head of Alignment, OpenAICo-led the Superalignment project with Ilya Sutskever from June 2023 until his resignation.Source →
- May 2024 – presentLead, Alignment Science team, AnthropicTeam focused on scalable oversight, weak-to-strong generalization and automated alignment research.Source →
On the record
22 Jan 2026 · Post · Musings on the Alignment Problem (Substack)
Once our models are so capable that on many tasks we don't understand their actions anymore, a lot of current approaches won't work the same way.
Lead, Alignment Science team, Anthropic at the time
Alignment is not solved but it increasingly looks solvable →
17 May 2024 · Post · X
safety culture and processes have taken a backseat to shiny products
Lead, Alignment Science team, Anthropic at the time
From the X thread announcing his resignation from OpenAI, as quoted by MIT Technology Review.