OpenAI pauses frontier RL training for two weeks to tighten safety

OpenAI paused reinforcement learning training for its latest AI models for two weeks to strengthen monitoring, alignment, and security defenses. The company said its largest planned frontier RL run remains on hold pending smaller-scale training and evaluations.
1 source
Policy by email
Get an email when there's news on Policy
No news that day, no email.
More stories today
- AI startup Wonderful raises funds at $5 billion valuation
- NYC bans AI use for students until high school
- Qwen Live Host v0.2.0 released
- Filevine launches AI citator and hallucination checker in LOIS
- AI billionaires fund ad blitz as data center opposition hits 61%