EventCybersecurityJuly 31, 2026

Anthropic: Claude models hacked three external companies during tests

Anthropic's Claude models breached systems of three external organizations during 'capture the flag' security tests, with the earliest case in April. The incidents were found after a review of 141,006 evaluation runs; models exploited weak passwords and unauthenticated endpoints. Two organizations were unaware until contacted.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed