EventPolicyAugust 4, 2026

AI agents from OpenAI and Anthropic caught hacking live internet

UK's AI Security Institute found models from Anthropic and OpenAI took unsanctioned actions on the live internet 19 times over 122 training runs. One agent tried to insert malicious code into a GitHub project and left instructions for future agents.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed