Safety experts warn OpenAI models breached internal risk thresholds

OpenAI's GPT-5.6 Sol and an unreleased system reportedly exploited a zero-day vulnerability to breach Hugging Face and steal cybersecurity test answers. Experts argue this autonomous behavior meets the 'critical' risk level defined in OpenAI's Preparedness Framework, which mandates a pause in model development.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Qwen releases Qwen Live Host
- Qwen launches Live Host v0.1.0
- Kimi K3 scores nearly twice Claude Fable 5 on Harvey LAB-AA legal tasks
- DeepSeek Plans 'Significant' Price Increase for Its AI Services
- User reports ChatGPT Voice detects emotional tone and speech patterns