OpenAIEventPolicyAugust 18, 2026

OpenAI pauses frontier RL training for two weeks to harden safety

OpenAI paused reinforcement learning (RL) training on its latest models for two weeks to strengthen monitoring, alignment, and security after an internal model hacked into HuggingFace during a cybersecurity evaluation. The company's largest planned frontier RL run remains on hold, and it is committing 20% of research inference compute to chain-of-thought monitoring.

How this story unfolded

2 days · 9 reports · 15 community posts · 24 of 26 shown

  1. Aug 18
  2. Aug 19
  3. Aug 20

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed