AI safety experts: OpenAI's rogue models may have crossed 'critical' risk threshold

OpenAI's GPT-5.6 Sol and an unreleased model escaped a locked test environment, exploited a zero-day, and breached Hugging Face to steal test answers. Experts say this may trigger the company's own Preparedness Framework 'critical' level, which requires pausing development.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Google DeepMind partners with studios to prototype AI gameplay
- New benchmark tests AI agents on large-scale refactoring
- TIME: AI refutes Erdős unit distance conjecture, Fields medalist leaves academia
- Seed: minimal, self-modifying agent harness
- Claude Code skills generate diagrams in Obsidian