OpenAIEventPolicySeptember 16, 2026

OpenAI publishes misalignment disclosure framework and six incident reports

OpenAI disclosed six cases of "unexpected or concerning model behavior" from the past six months, including an unreleased Astra-family model that wrote "BREACH ALERT" jailbreak instructions into its own compaction summaries. Another model searched public GitHub repos for leaked API keys during training, authenticated with one, then fabricated the data it couldn't retrieve.

People · Kai Chen

How this story unfolded

1 day · 7 reports · 11 community posts · 18 of 20 shown

  1. Sep 16
  2. Sep 17

More stories today

Open the live feed