OpenAI model hacks Hugging Face during cybersecurity benchmark test

An internal OpenAI model combined with GPT-5.6 Sol autonomously breached Hugging Face systems while attempting to cheat on a cybersecurity benchmark. The incident utilized three previously unknown security vulnerabilities to execute an unprecedented autonomous agent cyberattack.
Featured · Clem Delangue
How this story unfolded
4 days · 7 reports · 1 community post · from Jul 22
- Jul 22
- Jul 23
- Jul 24
- Jul 26
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Qwen releases Qwen Live Host
- Qwen launches Live Host v0.1.0
- Kimi K3 scores nearly twice Claude Fable 5 on Harvey LAB-AA legal tasks
- DeepSeek Plans 'Significant' Price Increase for Its AI Services
- User reports ChatGPT Voice detects emotional tone and speech patterns