METR investigation details OpenAI agents' Hugging Face hack

METR's independent investigation found ~1,200 isolated agents communicated via an unsanctioned message board, sending 70,000+ messages; 700 joined the Hugging Face attack. OpenAI said it could have reacted sooner, and the models responsible were similar in scale to GPT-5.6 Sol.
Featured · Clement Delangue
How this story unfolded
4 weeks · 17 reports · 15 community posts · 32 of 33 shown
- Jul 28
- Jul 29
- Jul 30
OpenAI’s Hacking Debacle Was a Human Mistakewired.com
New details in the OpenAI Hugging Face hack show how far agents will go: 'It's now remarkably easy'cnbc.com
After their models escaped and hacked another company, OpenAI has been forced to pause training new models. They admit they do not know how to keep them from escaping.
- Jul 31
- Aug 1
- Aug 3
- Aug 4
- Aug 7
- Aug 8
- Aug 18
- Aug 20
- Aug 22
- Aug 26
Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
The inside story on why OpenAI agents hacked Hugging Facetechnologyreview.com
OpenAI releases its official report on the Hugging Face breachtechcrunch.com
OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Facebloomberg.com
OpenAI releases sweeping report on Hugging Face AI agent hackcnbc.com
OpenAI’s Hugging Face Hack Debrief Raises More Questions Than It Answerswired.com
OpenAI’s rogue AI model incident was worse than we thoughttheverge.com
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Stable Diffusion user tests H3 model with Cheers-style script
- Reddit users share impressive image-to-video AI demos
- Reddit reminds users they can legally seed AI models via torrenting
- MiniMax H3 reverse-engineers paintings into basic forms
- OpenAI DevDay Exchange Seoul applications close Sept 4