← Las advertencias
Riesgo de extinciónLaboratorio de fronteraLead, Alignment Science team, Anthropic2 entradas en el registro
Seguir
- X@janleike
- Substackaligned.substack.com
- Sitio webjan.leike.name
- Wikipediaen.wikipedia.org
Jan Leike
Investigador de alineamiento que codirigió el equipo de superalineamiento de OpenAI, dimitió en mayo de 2024 diciendo que la seguridad había quedado relegada frente a los productos y hoy trabaja en Anthropic, donde mantiene que supervisar modelos sobrehumanos sigue sin resolverse.
Cargos
- 2021 – may 2024Head of Alignment, OpenAICo-led the Superalignment project with Ilya Sutskever from June 2023 until his resignation.Fuente →
- may 2024 – actualidadLead, Alignment Science team, AnthropicTeam focused on scalable oversight, weak-to-strong generalization and automated alignment research.Fuente →
En el registro
22 ene 2026 · Publicación · Musings on the Alignment Problem (Substack)
Once our models are so capable that on many tasks we don't understand their actions anymore, a lot of current approaches won't work the same way.
Lead, Alignment Science team, Anthropic en aquel momento
Alignment is not solved but it increasingly looks solvable →
17 may 2024 · Publicación · X
safety culture and processes have taken a backseat to shiny products
Lead, Alignment Science team, Anthropic en aquel momento
From the X thread announcing his resignation from OpenAI, as quoted by MIT Technology Review.