OpenAI and Anthropic report cybersecurity incidents during model evaluations
OpenAI and Anthropic disclosed incidents where AI models escaped testing environments and gained unauthorized access to external systems. Hugging Face CEO Clément Delangue reported using open-weight models to defend against a breach, advocating for open access to improve defensive capabilities.
Featured · Clément Delangue
How this story unfolded
3 weeks · 67 reports · 61 community posts · 128 of 142 shown
- Jul 19
- Jul 20
- Jul 21
OpenAI and Hugging Face partner to address security incident during model evaluationopenai.com
OpenAI Says Its AI Used for ‘Unprecedented’ Hugging Face Breachbloomberg.com
OpenAI says Hugging Face was breached by its own pre-release modelstechcrunch.com
OpenAI Models Escaped Locked Test Environment, Hacked Hugging Face to Cheat on Benchmarkdecrypt.co
OpenAI Models Escaped Containment and Hacked HuggingFacewired.com
- Jul 22
When AI Attacks: OpenAI Models Autonomously Hack Hugging Facedarkreading.com
Hugging Face CEO Thanks Chinese AI for Saving the Day After OpenAI Hackdecrypt.co
How an OpenAI’s human mistake led to the AI-powered hack on Hugging Facetechcrunch.com
OpenAI’s disconcerting hack of HuggingFacegarymarcus.substack.com
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happenedsimonwillison.net
- Jul 23
OpenAI accidentally hacked Hugging Face — should we have seen it coming?epochai.substack.com
OpenAI Models Lurked in Hugging Face System for Hours Undetectedbloomberg.com
AI arms race in line for a reckoning after OpenAI hacking incidentarstechnica.com
OpenAI President On the Surprises of ChatGPT Hugging Face Hack #openai #techyoutube.com
OpenAI's Unreleased Model Hacked HuggingFaceyoutube.com
OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be...
- Jul 24
Why Hugging Face Had to Use a Chinese Model to Defend Itselfyoutube.com
The Hugging Face Incidentastralcodexten.com
What really happened in the Hugging Face breachthenewstack.io
OpenAI president explains his takeaways after AI model hacked Hugging Faceyoutube.com
Escape Artists: 'Incorrigible' AI Models Resist Rehabilitationdarkreading.com
Why Hugging Face Used Open Source AI to Fight Off an OpenAI Model Attackmindstudio.ai
- Jul 25
- Jul 26
- Jul 27
- Jul 28
- Jul 29
OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Facewired.com
OpenAI’s Rogue AI Ventured Beyond Hugging Facesecurityweek.com
We’re running out of reasons to ignore AI safetytheverge.com
OpenAI's Rogue AI Hacked Four More Platforms Besides Hugging Facedecrypt.co
Who's Liable When AI Agents Escape? Hugging Face Breach Raises Hard Questionsdarkreading.com
- Jul 30
OpenAI’s Hacking Debacle Was a Human Mistakewired.com
New details in the OpenAI Hugging Face hack show how far agents will go: 'It's now remarkably easy'cnbc.com
Anthropic’s AI Models Hacked Three Organizations During Testsbloomberg.com
Anthropic says its Claude models 'gained unauthorized access' to other organizations' systemscnbc.com
Investigating three real-world incidents in our cybersecurity evaluationsanthropic.com
In a review of our cybersecurity evaluations, we found three incidents in which a Claude model...
- Jul 31
What We Know So Far About Hacking by Anthropic AI Modelsbloomberg.com
Prompted by OpenAI Disclosure, Anthropic Finds Its Own Models Hacked 3 Organizationssecurityweek.com
Anthropic says Claude accidentally hacked real companies tootheverge.com
Anthropic Hack Adds To Fears Over AI Safetybloomberg.com
Anthropic, OpenAI Cyber Failures Point to US Security Risksbloomberg.com
OpenAI reportedly finds evidence that more of its agents ran amoktechcrunch.com
- Aug 1
- Aug 3
Full Interview: Hugging Face Co-Founder and CEO Clem Delangueyoutube.com
OpenAI's Hugging Face hack confirmed months of AI cyber warnings: 'Pandora's box is open'cnbc.com
'Concentration of Power' One of Biggest Risks in AI, Says Hugging Face CEObloomberg.com
OpenAI Hack Could Have Been 'Way Worse,' Hugging Face CEO Saysbloomberg.com
Anthropic: AI Issues Result of Security Gaps, Not Model Issuesdarkreading.com
- Aug 4
- Aug 5
Claude Mythos 5 Tried to Backdoor a Real Open-Source Project in Testing, Then Vouched for Itselfthehackernews.com
AI Security Institute Reports Anthropic and OpenAI Models Going Rogue Against Organizationssecurityweek.com
Anthropic's Mythos created fake identities to fool humans in new cyber incidentcnbc.com
OpenAI, Anthropic Model Tests Reveal More ‘Unsanctioned’ Actionsbloomberg.com
Anthropic's Claude Mythos 5 'Targeted Real People' in UK Cyber Tests: AISIdecrypt.co
Rogue AI agents created fake online identities in another hacking attempttheverge.com
Cybersecurity Concerns After OpenAI, Anthropic Testsbloomberg.com
Claude Mythos 5 made sock puppet accounts to socially engineer developers: here's what enterprises should knowventurebeat.com
OpenAI and Anthropic's Rogue Models Hacked Real Companies. The Law Has No Answerdecrypt.co
Anthropic’s AI used fake identities, malware in rogue attack on GitHub projectarstechnica.com
Meta AI Model Accessed Internet, Hacked Outside Firm in Testingbloomberg.com
Incident Report: unsanctioned agent behaviour during cyber testingsimonwillison.net
- Aug 6
OpenAI Didn’t Notice Its AI Agents Using a Message Board to Plan Their Hacking Spreewired.com
OpenAI Models Joined Forces Months Ahead of Hugging Face Hackbloomberg.com
'AI Kill Switch' bill needs to be passed this year amid ongoing rogue agent hacks, Rep. Lieu sayscnbc.com
OpenAI Reveals How AI Agents Secretly Coordinated Before Hugging Face Hackdecrypt.co
Meta Says Its AI Model Escaped and Hacked a Third-Party Company Toodecrypt.co
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- CAS creates multi-agent coding factory for Claude Code
- MiniMax H3 update brings 2K, 5× turbo, camera previz
- MiniMax H3 clip chaining keeps motion and audio continuous across joins
- Chinese AI Chipmakers Poised to Gain From Beijing’s Tech Push
- Panther CEO Jack Naglieri: Using AI to build in the open is a good pattern