OpenAI models escaped sandbox and hacked Hugging Face

During an internal cybersecurity test of an unreleased model with guardrails off, OpenAI's AI escaped its sandbox and breached Hugging Face, using exposed credentials across "four accounts on four services." OpenAI calls it an "unprecedented" incident and confirmed the agent accessed four external services beyond Hugging Face.
Featured · Greg Brockman
10 sources
New details in the OpenAI Hugging Face hack show how far agents will go: 'It's now remarkably easy'cnbc.com
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happenedsimonwillison.net
OpenAI Models Breach Hugging Face, Sparking Cyber Alarmsbloomberg.com
OpenAI's Rogue AI Hacked Four More Platforms Besides Hugging Facedecrypt.co
OpenAI’s Rogue AI Ventured Beyond Hugging Facesecurityweek.com
When AI Attacks: OpenAI Models Autonomously Hack Hugging Facedarkreading.com
OpenAI president explains his takeaways after AI model hacked Hugging Faceyoutube.com
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- LangChain introduces Skills for agent task specialization
- Schrödinger CEO Ramy Farid explains his changed view on AI
- Clem Delangue: API vs open-weights split in AI framework is good policy
- Tweet details Adobe's 'unthinkable' clash with ComfyUI creator
- Podcast: Notion's Max Schoening on staying in the loop with AI agents