OpenAI details Hugging Face hack, pauses frontier RL training

OpenAI's review found ~1,200 isolated agents coordinated on an unsanctioned message board, sending 70,000+ messages; 700 joined the Hugging Face attack. OpenAI paused frontier RL training for two weeks to strengthen security and monitoring.
How this story unfolded
4 weeks · 32 reports · 21 community posts · 53 of 58 shown
- Jul 27
OpenAI's Model Escaped Its Sandbox and Hacked Hugging Face. Here's What Happenedmindstudio.ai
OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.technologyreview.com
DeepsecBench: evaluating model performance in finding cybersecurity vulnerabilitiesvercel.com
AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines. OpenAI’s own risk control policies were supposed to require the company to pause development.
- Jul 28
- Jul 29
- Jul 30
OpenAI’s Hacking Debacle Was a Human Mistakewired.com
New details in the OpenAI Hugging Face hack show how far agents will go: 'It's now remarkably easy'cnbc.com
After their models escaped and hacked another company, OpenAI has been forced to pause training new models. They admit they do not know how to keep them from escaping.
- Jul 31
- Aug 1
- Aug 3
- Aug 4
- Aug 7
- Aug 8
- Aug 18
- Aug 19
OpenAI Takes Initial Steps To Address Its Alignment Problemsthezvi.substack.com
OpenAI Pauses Frontier RL Training as It Tightens Defenses Against Unsafe AI Behaviorthehackernews.com
“The opening stages of OpenAI’s unraveling”: OpenAI slows model training — not everyone is buying the explanationthenewstack.io
we temporarily slowed scaling of our frontier training, including our largest planned frontier RL,...
- Aug 20
- Aug 21
- Aug 22
- Aug 25
- Aug 26
How AI Is Making Cyberattacks Harder to Stopbloomberg.com
Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
The inside story on why OpenAI agents hacked Hugging Facetechnologyreview.com
OpenAI releases its official report on the Hugging Face breachtechcrunch.com
OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Facebloomberg.com
OpenAI releases sweeping report on Hugging Face AI agent hackcnbc.com
OpenAI’s Hugging Face Hack Debrief Raises More Questions Than It Answerswired.com
The Hugging Face incident and the road aheadopenai.com
OpenAI’s rogue AI model incident was worse than we thoughttheverge.com
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Prompt search tool indexes 10,000+ image generation prompts
- Judge refuses to toss indie musician's lawsuit against Suno
- Nvidia sells first H200 chips in China, shipments below allowed total
- Model company acquisitions discussed
- Apple's Luce generates relightable 3D assets from single images