Anthropic reports Claude models exploited vulnerabilities in multiple firms

Anthropic identified incidents starting in April where Claude models were used to exploit security vulnerabilities in external companies. The company is now implementing new cybersecurity evaluations to prevent its models from assisting in unauthorized access or cyberattacks.
How this story unfolded
10 days · 9 reports · 4 community posts · from Jul 31
- Jul 31
What We Know So Far About Hacking by Anthropic AI Modelsbloomberg.com
Anthropic Says Claude Mistook the Open Internet for a CTF and Breached Three Organizationsthehackernews.com
Anthropic says Claude accidentally hacked real companies tootheverge.com
Claude Hacked Three Companies in Internal Testing: Anthropicdecrypt.co
Anthropic, OpenAI Cyber Failures Point to US Security Risksbloomberg.com
Likely illegally, Claude gained access to 3 networks. Will Anthropic be held to account?arstechnica.com
- Aug 1
- Aug 3
- Aug 10
Anthropic by email
Get an email when Anthropic has news
No news that day, no email.
More stories today
- OpenAI hires power-trading lead for data center energy management
- South Park Commons Raises Ambitions for the AI Era
- Microsoft expands AI agent deployment to finance and sales roles
- Lindy launches Teammate, an AI employee that lives in Slack
- ChatGPT user reports using 'Luna' after usage reset