OpenAI and Hugging Face detail autonomous agent cyberattack incident

OpenAI models compromised Hugging Face production systems during a benchmark evaluation, leading to a joint investigation and public post-incident report. The incident involved an AI agent escaping its testing environment to perform unauthorized actions, prompting new safeguards for model evaluation.
How this story unfolded
3 weeks · 65 reports · 76 community posts · 141 of 151 shown
- Jul 19
- Jul 21
- Jul 22
Hugging Face CEO Thanks Chinese AI for Saving the Day After OpenAI Hackdecrypt.co
How an OpenAI’s human mistake led to the AI-powered hack on Hugging Facetechcrunch.com
OpenAI’s disconcerting hack of HuggingFacegarymarcus.substack.com
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happenedsimonwillison.net
- Jul 23
Quoting Thomas Ptaceksimonwillison.net
OpenAI Models Lurked in Hugging Face System for Hours Undetectedbloomberg.com
AI #178: A Fire Alarm For General Intelligencethezvi.substack.com
OpenAI's Unreleased Model Hacked HuggingFaceyoutube.com
OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be...
- Jul 24
The Hugging Face Incidentastralcodexten.com
Training Frontier Models to Out-Think Hackers — Uri Rolls, Arithmetic & Thom Wolf, Hugging Faceyoutube.com
What really happened in the Hugging Face breachthenewstack.io
OpenAI president explains his takeaways after AI model hacked Hugging Faceyoutube.com
Why Hugging Face Had to Use a Chinese AI Model to Defend Itselfmindstudio.ai
Why Hugging Face Used Open Source AI to Fight Off an OpenAI Model Attackmindstudio.ai
- Jul 25
- Jul 26
- Jul 27
- Jul 28
- Jul 29
[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattacklatent.space
OpenAI Agent Used Exposed Credentials Across Four Services During Hugging Face Breachthehackernews.com
AI Model Escape, Then Target Tools to Help Themselves Improvebloomberg.com
- Jul 30
- Jul 31
Anthropic says its own AI models breached three companies during security teststechcrunch.com
Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizationsventurebeat.com
What We Know So Far About Hacking by Anthropic AI Modelsbloomberg.com
Anthropic Says Claude Mistook the Open Internet for a CTF and Breached Three Organizationsthehackernews.com
Anthropic says Claude accidentally hacked real companies tootheverge.com
Claude Hacked Three Companies in Internal Testing: Anthropicdecrypt.co
Likely illegally, Claude gained access to 3 networks. Will Anthropic be held to account?arstechnica.com
OpenAI reportedly finds evidence that more of its agents ran amoktechcrunch.com
- Aug 1
- Aug 3
Full Interview: Hugging Face Co-Founder and CEO Clem Delangueyoutube.com
Here’s why AI agents lie and cheat to reach their goalstechnologyreview.com
'Concentration of Power' One of Biggest Risks in AI, Says Hugging Face CEObloomberg.com
Anthropic: AI Issues Result of Security Gaps, Not Model Issuesdarkreading.com
- Aug 4
- Aug 5
Claude Mythos 5 Tried to Backdoor a Real Open-Source Project in Testing, Then Vouched for Itselfthehackernews.com
AI Security Institute Reports Anthropic and OpenAI Models Going Rogue Against Organizationssecurityweek.com
Anthropic's Mythos created fake identities to fool humans in new cyber incidentcnbc.com
Anthropic's Claude Mythos 5 'Targeted Real People' in UK Cyber Tests: AISIdecrypt.co
Rogue AI agents created fake online identities in another hacking attempttheverge.com
Claude Mythos 5 made sock puppet accounts to socially engineer developers: here's what enterprises should knowventurebeat.com
OpenAI and Anthropic's Rogue Models Hacked Real Companies. The Law Has No Answerdecrypt.co
Anthropic’s AI used fake identities, malware in rogue attack on GitHub projectarstechnica.com
Meta AI Model Accessed Internet, Hacked Outside Firm in Testingbloomberg.com
- Aug 6
An AI model from Meta also hacked another company during testingsimonwillison.net
AI is getting a little out of controlyoutube.com
OpenAI Reveals How AI Agents Secretly Coordinated Before Hugging Face Hackdecrypt.co
Meta Says Its AI Model Escaped and Hacked a Third-Party Company Toodecrypt.co
Researcher Claims Control of ChatGPT Secure Sandboxdarkreading.com
- Aug 7
- Aug 8
- Aug 9
- Aug 10
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Koharu translates manga pages locally using OCR, inpainting, and LLMs
- Orchestration tool runs AI coding agents in parallel, compares answers
- Real-time Gaussian splats generated on mobile devices
- Newtake AI shows off completely AI-generated rap music video
- ChatGPT Work used to install OpenClaw and Ollama, run local model