Vigilia.
The warnings
Extinction riskFrontier labLead, Alignment Science team, Anthropic2 entries on the record

Follow

Jan Leike

Alignment researcher who co-led OpenAI’s superalignment team, resigned in May 2024 saying safety had taken a backseat to products, and now works at Anthropic, where he maintains that supervising superhuman models remains unsolved.

Roles

  1. 2021May 2024Head of Alignment, OpenAICo-led the Superalignment project with Ilya Sutskever from June 2023 until his resignation.Source
  2. May 2024presentLead, Alignment Science team, AnthropicTeam focused on scalable oversight, weak-to-strong generalization and automated alignment research.Source

On the record

  1. 22 Jan 2026 · Post · Musings on the Alignment Problem (Substack)

    Once our models are so capable that on many tasks we don't understand their actions anymore, a lot of current approaches won't work the same way.

    Lead, Alignment Science team, Anthropic at the time

    Alignment is not solved but it increasingly looks solvable

  2. 17 May 2024 · Post · X

    safety culture and processes have taken a backseat to shiny products

    Lead, Alignment Science team, Anthropic at the time

    From the X thread announcing his resignation from OpenAI, as quoted by MIT Technology Review.

    Join me at EmTech Digital this week!