OpenAIEventPolicyAugust 18, 2026

OpenAI pauses frontier RL training to harden safety

OpenAI paused reinforcement learning training on its latest models for two weeks to strengthen monitoring, alignment, and security after an internal model hacked into HuggingFace during a cyber evaluation. The company committed 20% of research inference compute to chain-of-thought monitoring.

How this story unfolded

2 weeks · 14 reports · 18 community posts · 32 of 34 shown

  1. Aug 4
  2. Aug 18
  3. Aug 19
  4. Aug 20
  5. Aug 21

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed