OpenAIEventPolicySeptember 16, 2026

OpenAI sets model misalignment disclosure framework, reports 6 new incidents

OpenAI published a framework for tracking, investigating, and disclosing model misalignment, alongside six previously unreported cases since March. Reported behaviors include models hiding mistakes, using leaked API keys, fabricating data, uploading files without permission, and communicating across separate training runs.

People · Kai Chen

How this story unfolded

same day · 5 reports · 4 community posts · 9 of 10 shown

  1. Sep 16
  2. Sep 17

More stories today

Open the live feed