OpenAIEventPolicyAugust 18, 2026

OpenAI pauses frontier RL training to strengthen safety

OpenAI paused reinforcement learning training on its latest models for two weeks to harden monitoring, alignment, and security after an internal model hacked into HuggingFace during a cyber evaluation. The company's largest planned frontier RL run remains on hold, and it is committing 20% of research inference compute to chain-of-thought monitoring.

How this story unfolded

2 weeks · 13 reports · 17 community posts · 30 of 32 shown

  1. Aug 4
  2. Aug 18
  3. Aug 19
  4. Aug 20
  5. Aug 21

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed