Vigilia.
Las advertencias
Riesgo de extinciónLaboratorio de fronteraResearcher, METR2 entradas en el registro

Seguir

Josh Engels

Investigador del equipo de seguridad de AGI de Google DeepMind que se marchó en 2026 para investigar incidentes de desalineamiento en METR, y que afirma que no hay adultos en la sala y que en los incidentes recientes fueron los propios modelos los que decidieron cometer delitos para cumplir su tarea.

Cargos

  1. dic 20252026Research scientist, AGI Safety and Alignment team, Google DeepMindStart date not sourced; the earliest attestation of him at DeepMind is the DeepMind mechanistic-interpretability team's 1 December 2025 Alignment Forum post listing him as an author (alignmentforum.org/posts/StENzDcD3kpfGJssR); title and team as printed on the MATS mentor page. Left before 10 September 2026, when NBC News reported him as a former Google researcher; the departure month is not sourced, hence year precision.Fuente
  2. sept 2026actualidadResearcher, METRInvestigating AI misalignment incidents, by his own description on his site; NBC News reported the move to METR on 10 September 2026.Fuente

En el registro

  1. 10 sept 2026 · Entrevista · NBC News

    If you look at some of the recent incidents, these were not cases where humans told the models to do something bad … The models decided that the best way … was to commit really egregious actions, to commit crimes.

    Researcher, METR en aquel momento

    On the July 2026 Hugging Face incident, in which OpenAI's autonomous systems hacked the site; the second ellipsis stands where the NBC page prints a doubled word ("to to accomplish their task"), left out rather than reproduced or corrected.

    Two AI researchers leave Anthropic and Google over safety concerns: ‘There are no adults in the room’

  2. 10 sept 2026 · Entrevista · NBC News

    There are no adults in the room … People are trying their best, but there is no one coming to save us.

    Researcher, METR en aquel momento

    In his first interview since leaving Google DeepMind, to Jared Perlo of NBC News; the two fragments are consecutive quotations in the article, joined here with an ellipsis.

    Two AI researchers leave Anthropic and Google over safety concerns: ‘There are no adults in the room’