OpenAI pauses frontier RL training, adds security after Hugging Face breach

OpenAI paused reinforcement learning training for two weeks and halted its largest planned frontier RL run after internal evaluations showed its upcoming Astra model may reach 'critical' cybersecurity capability. New safeguards include sandboxing, network isolation, and a monitoring system that pages teams within 30 minutes, consuming ~20% of inference compute.
Featured · Amelia Glaese
How this story unfolded
2 weeks · 7 reports · 1 community post · from Aug 4
- Aug 4
- Aug 18
- Aug 19
- Aug 20
- Aug 21
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- AWS Quick and fal enable agentic creative workflows
- Anthropic opens 10,000 free Claude seats for scientists
- Researcher breaks Claude Code Opus 5 auto mode with 80% success
- Nvidia CEO Jensen Huang: I wish I had invested more in AI frontier labs
- Apple introduces rubric-based alignment for grounded QA