Vigilia.
Gli avvertimenti
Rischio di estinzioneLaboratorio di frontieraResearcher, METR2 voci a registro

Seguire

Josh Engels

Ricercatore del team di sicurezza AGI di Google DeepMind, uscito nel 2026 per indagare sugli incidenti di disallineamento in METR, che afferma che non ci sono adulti nella stanza e che negli incidenti recenti sono stati i modelli stessi a scegliere di commettere reati per portare a termine il proprio compito.

Ruoli

  1. dic 20252026Research scientist, AGI Safety and Alignment team, Google DeepMindStart date not sourced; the earliest attestation of him at DeepMind is the DeepMind mechanistic-interpretability team's 1 December 2025 Alignment Forum post listing him as an author (alignmentforum.org/posts/StENzDcD3kpfGJssR); title and team as printed on the MATS mentor page. Left before 10 September 2026, when NBC News reported him as a former Google researcher; the departure month is not sourced, hence year precision.Fonte
  2. set 2026oggiResearcher, METRInvestigating AI misalignment incidents, by his own description on his site; NBC News reported the move to METR on 10 September 2026.Fonte

A registro

  1. 10 set 2026 · Intervista · NBC News

    If you look at some of the recent incidents, these were not cases where humans told the models to do something bad … The models decided that the best way … was to commit really egregious actions, to commit crimes.

    Researcher, METR all’epoca

    On the July 2026 Hugging Face incident, in which OpenAI's autonomous systems hacked the site; the second ellipsis stands where the NBC page prints a doubled word ("to to accomplish their task"), left out rather than reproduced or corrected.

    Two AI researchers leave Anthropic and Google over safety concerns: ‘There are no adults in the room’

  2. 10 set 2026 · Intervista · NBC News

    There are no adults in the room … People are trying their best, but there is no one coming to save us.

    Researcher, METR all’epoca

    In his first interview since leaving Google DeepMind, to Jared Perlo of NBC News; the two fragments are consecutive quotations in the article, joined here with an ellipsis.

    Two AI researchers leave Anthropic and Google over safety concerns: ‘There are no adults in the room’