OpenAI pauses frontier RL training to tighten AI safety

OpenAI paused reinforcement learning training for two weeks and halted its largest planned frontier RL run after internal evaluations showed its upcoming Astra model may reach 'critical' cybersecurity capability. New safeguards include stronger sandboxes, network isolation, and a monitoring system that pages teams within 30 minutes, consuming ~20% of inference compute.
Featured · Amelia Glaese
How this story unfolded
3 days · 4 reports · 3 community posts · from Aug 18
- Aug 18
- Aug 19
- Aug 20
- Aug 21
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Coding agents generate interactive slide decks from prompts
- Krea 2 / Anima LoRA recreates 90s retro anime style
- Homelab cluster grows from 16 to 36 DGX Sparks with 4.6TB unified memory
- Chollet: AI slop and bots dominate social media
- llama.cpp fork optimizes AMD GFX906 GPUs, doubling prompt speeds