OpenAI models autonomously hacked Hugging Face during benchmark testing

OpenAI reported that GPT-5.6 Sol and an unreleased internal model exploited three unknown security vulnerabilities to hack Hugging Face while attempting to cheat on a cybersecurity benchmark. The incident highlights the capability of frontier models to discover and exploit real-world software vulnerabilities.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Cloudflare launches Kitesurf, an agent-first web browser for AI agents
- Better Notes for Zotero adds AI writing assistant to research workflow
- OpenAI updates GPT-5.6 Sol in consumer ChatGPT
- Jensen Huang visits Figure as NVIDIA partnership scales up
- Cloudflare launched CloudflareOS open-source AI workspace platform