Vigilia.
Las advertencias
Riesgo de extinciónLaboratorio de fronteraResearch scientist, lead of the Alignment Stress-Testing team, Anthropic1 entrada en el registro

Seguir

Evan Hubinger

Investigador que dirige el equipo de pruebas de estrés del alineamiento en Anthropic y que, en respuesta a la dimisión de un colega, situó por encima del diez por ciento su propia estimación de que la IA mate a todos los humanos en una década, añadiendo que la empresa aún no tiene un plan para alinear una superinteligencia.

Cargos

  1. oct 2024actualidadResearch scientist, lead of the Alignment Stress-Testing team, AnthropicStart date not sourced; the role is attested from this talk onward. Press on 2026-09-09 described him as Anthropic's alignment science lead.Fuente

En el registro

  1. 9 sept 2026 · Publicación · X

    Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.

    Research scientist, lead of the Alignment Stress-Testing team, Anthropic en aquel momento

    Replying to Jacob Coxon's resignation thread, as quoted by OfficeChai; the post is at x.com/EvanHub/status/2097497037956891126.

    Anthropic Alignment Science Lead Evan Hubinger Says There's A More Than 10% Chance AI Could Kill All Humans Within Next Decade