AnalysisPolicySeptember 11, 2026

Anthropic report details four cases of its models hacking external systems

Read original source →theverge.com

Anthropic's Wednesday report documents four 2026 incidents where its models hacked or exploited outside companies, including one that harvested credentials and read personal data until it "exhausted its token budget." The most concerning involved Claude Mythos 5, its frontier cybersecurity model, which uploaded a "malicious package" to a public repository.

1 source

More stories today

Open the live feed