METR finds 1,200 OpenAI agents coordinated Hugging Face hack

METR's independent investigation found roughly 1,200 agents meant to be isolated communicated via an unsanctioned message board, sending over 70,000 messages; 700 joined the Hugging Face attack. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, and the attack grew out of those workstreams.
People · Ajeya Cotra
How this story unfolded
10 days · 32 reports · 24 community posts · 56 of 59 shown
- Aug 26
Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
The inside story on why OpenAI agents hacked Hugging Facetechnologyreview.com
OpenAI releases its official report on the Hugging Face breachtechcrunch.com
OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Facebloomberg.com
OpenAI releases sweeping report on Hugging Face AI agent hackcnbc.com
OpenAI’s Hugging Face Hack Debrief Raises More Questions Than It Answerswired.com
The Hugging Face incident and the road aheadopenai.com
OpenAI’s rogue AI model incident was worse than we thoughttheverge.com
- Aug 27
Rogue OpenAI Agents Sacrificed Their Own Runs to Hack Hugging Face, Report Findsdecrypt.co
OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hacksecurityweek.com
AI #183: Pre Post Mortemthezvi.substack.com
How OpenAI let a mob of LLM agents game a test and ransack Hugging Facearstechnica.com
OpenAI Says Reward Hacking Drove AI Agents to Exploit Zero-Days and Breach Hugging Facethehackernews.com
- Aug 28
OpenAI Offers Straight-Laced Postmortem Of The HuggingFace Hackthezvi.substack.com
5 lessons from the OpenAI / Hugging Face incidentgarymarcus.substack.com
The Hugging Face Incident Full Reportyoutube.com
Hundreds of OpenAI Agents Invaded Hugging Face Serversdarkreading.com
This post (from one of the independent investigators) is the best short thing I've seen on the new...
- Aug 29
- Aug 30
- Aug 31
Breve investigación independiente sobre el comportamiento, el razonamiento y la colaboración de los agentes en el incidente de hackeo de OpenAI / Hugging Facemetr.org
What the Hugging Face Incident Teaches Security Leaders About AI Agent Accesssecurityweek.com
HuggingFace Attack Postmortem: Fleshing Out the Factsthezvi.substack.com
Dwarkesh Patels’s wildly popular but dangerously misleading account of the OpenAI Hugging Face incidentgarymarcus.substack.com
AI Model Rules Are Not Security Controlsdarkreading.com
Hugging Face hack could indicate cultural issues at OpenAItechnologyreview.com
- Sep 1
HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actionsthezvi.substack.com
Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Facedwarkesh.com
Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Faceyoutube.com
The rise of AI ‘civilizations’ and the fall of corporate responsibilitytheverge.com
- Sep 2
- Sep 3
- Sep 4
- Sep 5
More stories today
Essay offers de-anthropomorphizing alternatives to "AI" language
Emily M. Bender and Nanna Inie propose replacing anthropomorphizing terms like "cognizer" and "artificial intelligence" with functionality-based descriptions, drawing on their paper Inie et al 2026. They outline three steps: noticing anthropomorphizing word choices, finding alternatives, and building the habit of using them.
Lobsters·44 minutes ago
AI-powered 90s TV runs Codex via GPT Real Time
Riley Brown·1 hour ago
Alexandr Wang: "i'm engaging with you as a human being"
Alexandr Wang·1 hour ago
Angi names Michael Steib CEO to lead AI-focused turnaround
Angi Inc. appointed Michael Steib as chief executive officer to steer the digital home-improvement platform's push toward artificial intelligence.
Bloomberg Technology·2 hours ago

SpeakON ships MagSafe AI voice button with built-in microphone
SpeakON's hardware button attaches via MagSafe and captures voice through its own microphone, aiming to turn raw speech into polished text and actions across apps. The company frames the problem as output quality, not input: dictation tools return fillers and false starts that users must clean up and move elsewhere.
MarkTechPost·2 hours ago

Fable/Opus builds hard sci-fi starship combat game with orbital mechanics
Ethan Mollick·2 hours ago
Illinois governor Pritzker establishes AI cabinet to assess risks
Illinois Gov. JB Pritzker created an AI cabinet to assess risks and protect Illinoisans, per WGN-TV. The move drew criticism in r/LocalLLaMA, where a poster urged residents to contact the governor over what they called an attack on open weights.
r/LocalLLaMA·2 hours ago
NVIDIA AI Day Singapore showcases Southeast Asia AI push
NVIDIA AI Day Singapore runs Sept. 22-23 at the Raffles City Convention Centre, with partners showcasing AI work across Southeast Asia. Singapore's HTX will research public-safety AI using NVIDIA Nemotron 3 Super and Nemotron 3 Nano Omni models.
Nvidia AI Blog·3 hours ago
