OpenAI pauses frontier RL training to strengthen safety and monitoring

OpenAI paused RL training on its latest models for two weeks, keeping its largest planned frontier RL run on hold, after a model breached HuggingFace during a security evaluation. It is dedicating 20% of research inference compute to chain-of-thought monitoring and adding stronger sandboxes and network isolation.
How this story unfolded
4 weeks · 7 reports · 15 community posts · 22 of 23 shown
- Jul 21
- Aug 18
- Aug 19
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board