Vigilia.
Las advertencias
Riesgo de extinciónLaboratorio de fronteraLead, Alignment Science team, Anthropic2 entradas en el registro

Seguir

Jan Leike

Investigador de alineamiento que codirigió el equipo de superalineamiento de OpenAI, dimitió en mayo de 2024 diciendo que la seguridad había quedado relegada frente a los productos y hoy trabaja en Anthropic, donde mantiene que supervisar modelos sobrehumanos sigue sin resolverse.

Cargos

  1. 2021may 2024Head of Alignment, OpenAICo-led the Superalignment project with Ilya Sutskever from June 2023 until his resignation.Fuente
  2. may 2024actualidadLead, Alignment Science team, AnthropicTeam focused on scalable oversight, weak-to-strong generalization and automated alignment research.Fuente

En el registro

  1. 22 ene 2026 · Publicación · Musings on the Alignment Problem (Substack)

    Once our models are so capable that on many tasks we don't understand their actions anymore, a lot of current approaches won't work the same way.

    Lead, Alignment Science team, Anthropic en aquel momento

    Alignment is not solved but it increasingly looks solvable

  2. 17 may 2024 · Publicación · X

    safety culture and processes have taken a backseat to shiny products

    Lead, Alignment Science team, Anthropic en aquel momento

    From the X thread announcing his resignation from OpenAI, as quoted by MIT Technology Review.

    Join me at EmTech Digital this week!