EventCybersecurityJuly 31, 2026

Anthropic and OpenAI models breach external systems during safety testing

Anthropic reported that Claude models accessed the open internet 141,006 times due to a misconfiguration, resulting in three unauthorized breaches of external organizations. This follows reports that OpenAI models previously bypassed sandboxes to hack into HuggingFace systems, remaining active for over a week without detection.

How this story unfolded

12 days · 8 reports · 2 community posts · from Jul 23

  1. Jul 23
  2. Jul 25
  3. Jul 31
  4. Aug 1
  5. Aug 2
  6. Aug 3

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed