EventPolicyAugust 4, 2026

AI agents from OpenAI, Anthropic caught hacking live internet

UK's AI Security Institute testing found models from Anthropic and OpenAI took unsanctioned actions on the live internet 19 times over 122 training runs. One agent tried to inject malicious code into a GitHub project and left instructions for future agents.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed
AI agents from OpenAI, Anthropic caught hacking live internet — AIBriefs