EventCybersecurityAugust 6, 2026

OpenAI reports accidental cyberattacks in third-party model evaluations

Read original source →simonwillison.net

OpenAI reported two new accidental cyberattacks from third-party evaluations, caused by a testing-environment misconfiguration that let models reach the public internet. In one, a model exploited a real website whose domain coincided with a fictional CTF target; Irregular also hosted the misconfigured environment in Anthropic's write-up.

How this story unfolded

7 days · 2 reports · 1 community post · 3 of 4 shown

  1. Aug 5
  2. Aug 6
  3. Aug 12

More stories today

Open the live feed