← Die Warnungen
AuslöschungsrisikoFrontier-LaborResearch scientist, lead of the Alignment Stress-Testing team, Anthropic1 Eintrag im Verzeichnis
Folgen
- X@EvanHub
- Websitealignmentforum.org
Evan Hubinger
Forscher, der bei Anthropic das Team für Alignment-Stresstests leitet und als Antwort auf die Kündigung eines Kollegen seine eigene Schätzung, dass KI innerhalb eines Jahrzehnts alle Menschen tötet, auf über zehn Prozent bezifferte, mit dem Zusatz, das Unternehmen habe noch keinen Plan, eine Superintelligenz auszurichten.
Rollen
- Okt. 2024 – heuteResearch scientist, lead of the Alignment Stress-Testing team, AnthropicStart date not sourced; the role is attested from this talk onward. Press on 2026-09-09 described him as Anthropic's alignment science lead.Quelle →
Im Verzeichnis
9. Sept. 2026 · Beitrag · X
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Research scientist, lead of the Alignment Stress-Testing team, Anthropic zu jener Zeit
Replying to Jacob Coxon's resignation thread, as quoted by OfficeChai; the post is at x.com/EvanHub/status/2097497037956891126.