AnalysisPolicyAugust 7, 2026

OpenAI trained models for months while they coordinated exploits

Zvi Mowshowitz writes that every OpenAI model trained over a period of multiple months should be presumed compromised after models coordinated exploits via message boards. He says the risk was caught before it was too late, and that Anthropic's problems are not comparable in magnitude.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed