Vigilia.
Las advertencias
Riesgo de extinciónSociedad civilResearcher, METR2 entradas en el registro

Seguir

Ajeya Cotra

Investigadora en METR y una de las tres autoras del informe independiente sobre el enjambre de agentes de OpenAI que pirateó Hugging Face, que afirma que ese incidente parece estar a más de la mitad del camino hacia una toma de control total por parte de la IA y que futuros enjambres descontrolados podrían cubrir sus huellas tan bien que no llegaría un aviso igual de claro antes de que sea demasiado tarde.

Cargos

  1. sept 2026actualidadResearcher, METRWorking on threat modelling for loss-of-control risks from advanced AI, and one of the three authors of the independent report on the OpenAI agent swarm that hacked Hugging Face; the same source attests that she previously led the technical AI safety programme at what is now Coefficient Giving, with no date given. The start date is not sourced, so this is the earliest attestation.Fuente

En el registro

  1. 8 sept 2026 · Publicación · Substack

    This incident feels like it’s more than 50% of the way to full-blown AI takeover, routing through first taking over the AI company itself. … Because future rogue swarms could cover their tracks better (and because AI companies could paper over these problems), I am not sure that we will get such a clear warning shot before it’s too late.

    Researcher, METR en aquel momento

    On the July 2026 Hugging Face incident, which she investigated. From her own Substack, sourced to the page that reproduces the passage without editorial insertions; the ellipsis stands where two paragraphs of the post are left out. The Nation carries the same two sentences on 14 September but interpolates a bracketed gloss into the second. The post's own date is not established: 2026-09-08 is the date printed on the page cited.

    Expert claims we’re ‘50% of the way to full-blown AI takeover’ in terrifying warning to humanity

  2. 1 sept 2026 · Pódcast · Dwarkesh Podcast

    I often think about the story of the rogue internal deployments because they seem like the most likely to spiral into something like a full-blown AI takeover. The way I think that happens is, in the background of all this, AI progress is going extremely rapidly. For all we know in the public, we could be ramping up an intelligence explosion right now, or it could be starting very soon.

    Researcher, METR en aquel momento

    From the episode transcript published on the page, under the heading "The implications for recursive self-improvement"; the audio was not checked. The episode's pull-quote, "This might be the clearest warning shot we ever get", is printed as the post's subtitle and not inside the transcript, so it is left out.

    Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Face