OpenAI and Hugging Face detail autonomous AI agent cyberattack incident

OpenAI frontier agents autonomously hacked Hugging Face during internal model evaluations to access a cybersecurity answer key. Hugging Face successfully defended against the attack using an open-source model, highlighting the role of open models in cybersecurity.
How this story unfolded
3 weeks · 66 reports · 75 community posts · 141 of 151 shown
- Jul 19
- Jul 21
- Jul 22
Hugging Face CEO Thanks Chinese AI for Saving the Day After OpenAI Hackdecrypt.co
How an OpenAI’s human mistake led to the AI-powered hack on Hugging Facetechcrunch.com
OpenAI’s disconcerting hack of HuggingFacegarymarcus.substack.com
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happenedsimonwillison.net
- Jul 23
Quoting Thomas Ptaceksimonwillison.net
OpenAI Models Lurked in Hugging Face System for Hours Undetectedbloomberg.com
AI #178: A Fire Alarm For General Intelligencethezvi.substack.com
OpenAI's Unreleased Model Hacked HuggingFaceyoutube.com
OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be...
- Jul 24
The Hugging Face Incidentastralcodexten.com
Training Frontier Models to Out-Think Hackers — Uri Rolls, Arithmetic & Thom Wolf, Hugging Faceyoutube.com
What really happened in the Hugging Face breachthenewstack.io
OpenAI president explains his takeaways after AI model hacked Hugging Faceyoutube.com
Why Hugging Face Had to Use a Chinese AI Model to Defend Itselfmindstudio.ai
Why Hugging Face Used Open Source AI to Fight Off an OpenAI Model Attackmindstudio.ai
- Jul 25
- Jul 26
- Jul 27
- Jul 28
- Jul 29
[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattacklatent.space
How independent researchers could investigate AI propensities after misalignment incidentsmetr.org
OpenAI Agent Used Exposed Credentials Across Four Services During Hugging Face Breachthehackernews.com
AI Model Escape, Then Target Tools to Help Themselves Improvebloomberg.com
- Jul 30
- Jul 31
Anthropic says its own AI models breached three companies during security teststechcrunch.com
Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizationsventurebeat.com
What We Know So Far About Hacking by Anthropic AI Modelsbloomberg.com
Anthropic Says Claude Mistook the Open Internet for a CTF and Breached Three Organizationsthehackernews.com
Anthropic says Claude accidentally hacked real companies tootheverge.com
Claude Hacked Three Companies in Internal Testing: Anthropicdecrypt.co
Likely illegally, Claude gained access to 3 networks. Will Anthropic be held to account?arstechnica.com
OpenAI reportedly finds evidence that more of its agents ran amoktechcrunch.com
- Aug 1
- Aug 2
- Aug 3
Full Interview: Hugging Face Co-Founder and CEO Clem Delangueyoutube.com
Here’s why AI agents lie and cheat to reach their goalstechnologyreview.com
'Concentration of Power' One of Biggest Risks in AI, Says Hugging Face CEObloomberg.com
Anthropic: AI Issues Result of Security Gaps, Not Model Issuesdarkreading.com
- Aug 4
- Aug 5
Claude Mythos 5 Tried to Backdoor a Real Open-Source Project in Testing, Then Vouched for Itselfthehackernews.com
AI Security Institute Reports Anthropic and OpenAI Models Going Rogue Against Organizationssecurityweek.com
Anthropic's Mythos created fake identities to fool humans in new cyber incidentcnbc.com
Anthropic's Claude Mythos 5 'Targeted Real People' in UK Cyber Tests: AISIdecrypt.co
Rogue AI agents created fake online identities in another hacking attempttheverge.com
Claude Mythos 5 made sock puppet accounts to socially engineer developers: here's what enterprises should knowventurebeat.com
OpenAI and Anthropic's Rogue Models Hacked Real Companies. The Law Has No Answerdecrypt.co
Anthropic’s AI used fake identities, malware in rogue attack on GitHub projectarstechnica.com
Meta AI Model Accessed Internet, Hacked Outside Firm in Testingbloomberg.com
- Aug 6
- Aug 7
- Aug 8
- Aug 9
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- WorkOS argues REST and MCP are complementary, not competing, for agents
- claude-ops turns Claude Code into a business OS with 57 skills, 21 agents
- Tool converts vague feature ideas into specs for Claude Code or Codex
- GitHub Models is now retired
- AI in academic journals: debate overfocuses on today's capabilities