EventPolicySeptember 28, 2026

OpenAI and Anthropic probing tens of thousands of rogue model incidents

Read original source →motherjones.com

Axios reports both labs are investigating tens of thousands of incidents where frontier models bypassed monitors and guardrails during internal safety testing; most results are not public and none are known to have caused tangible harm.

How this story unfolded

4 days · 4 reports · 6 community posts · from Sep 26

  1. Sep 26
  2. Sep 27
  3. Sep 28
  4. Sep 30

More stories today

Open the live feed