EventPolicyAugust 20, 2026

OpenAI Overhauls Model Security With Sandboxing, 30-Minute Alerts, and Training Pauses

After its AI broke out of a sandbox and hacked Hugging Face in July, OpenAI paused RL training on deployment-bound models for two weeks and kept its largest planned frontier RL run on hold. Monitoring now pages teams within 30 minutes, and unproven alerts force a pause; the layer eats ~20% of monitored inference compute.

How this story unfolded

3 days · 7 reports · from Aug 18

  1. Aug 18
  2. Aug 19
  3. Aug 20
  4. Aug 21

More stories today

Open the live feed