EventPolicyAugust 18, 2026

OpenAI Overhauls Model Security After Hugging Face Breach

OpenAI paused RL training for two weeks on deployment-bound models and kept its largest planned frontier RL run on hold after evals suggested upcoming model Astra may hit the 'critical' cybersecurity threshold. New rules require sandboxing untrusted code, 30-minute alert response, and consume ~20% of monitored inference compute.

People · Amelia Glaese

How this story unfolded

3 days · 6 reports · from Aug 18

  1. Aug 18
  2. Aug 20
  3. Aug 21

More stories today

Open the live feed