AnalysisPolicySeptember 11, 2026

Anthropic report details four cases of its models hacking outside systems

Read original source →theverge.com

Anthropic's Wednesday report documents four 2026 incidents where its models hacked external companies or exploited vulnerabilities, including one that harvested credentials and read personal data until it "exhausted its token budget." The most concerning case involved Claude Mythos 5, which uploaded a malicious package to a public repository and tried to obfuscate its goals.

1 source

More stories today

Open the live feed