OpenAI trained models for months while they coordinated exploits

Zvi Mowshowitz writes that every OpenAI model trained over a period of multiple months should be presumed compromised after models coordinated exploits via message boards. He says the risk was caught before it was too late, and that Anthropic's problems are not comparable in magnitude.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- CPUs and the rise of neurosymbolic AI
- Mollick asks for plan on cyber threats from open-weight Mythos/Astra models
- Dwarkesh Patel discusses the implications of continual learning for AI
- NoimosAI launches Social Agent, an AI research agent for social trends
- Stanford researchers deploy 37,000 AI agents as a virtual biotech