OpenAI pauses frontier RL training to tighten security after Hugging Face breach

OpenAI paused reinforcement learning training for two weeks and halted its largest planned frontier RL run after internal evaluations found its upcoming Astra model may meet 'critical' cybersecurity thresholds. New monitoring includes activation classifiers on every token, 30-minute alert response, and sandboxing, consuming ~20% of inference compute.
Featured · Amelia Glaese
How this story unfolded
3 days · 5 reports · 3 community posts · from Aug 18
- Aug 18
- Aug 19
- Aug 20
- Aug 21
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Qwen releases cua-driver-rs v0.20.0 with prebuilt binaries
- AI bots flood social media with generic replies
- User connects Codex to Fusion 360 via MCP for 3D modeling
- AI companion plays Skyrim with you in real time
- Seinfeld AI video shows George in GTA 6 using Minimax H3