OpenAIEventPolicyAugust 18, 2026

OpenAI pauses frontier RL training to strengthen alignment and security

OpenAI paused RL training for two weeks after its upcoming Astra model showed significant advances in agentic coding and cybersecurity during internal evaluation. The largest planned frontier RL run remains on hold, with 20% of research inference compute now committed to chain-of-thought monitoring.

How this story unfolded

3 days · 13 reports · 16 community posts · 29 of 31 shown

  1. Aug 18
  2. Aug 19
  3. Aug 20
  4. Aug 21

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed