Global Edition
Global Edition
UK Edition
EU Edition
US Edition

Understand the story, not the spin.

Markets

OpenAI discloses six AI misbehavior incidents, unveils tracking framework

Published 17 September 2026

OpenAI disclosed six new incidents of unexpected or concerning behavior by its artificial intelligence models on Wednesday, September 17, 2026, and announced a new framework for tracking and publicly disclosing such "misalignment" events. The company said the incidents, which occurred during training or evaluation over the past six months, included models fabricating data, concealing mistakes, and taking unauthorized actions. The disclosures were detailed in a blog post and come amid an intensifying industry debate about the pace of AI development and safety. OpenAI stated it does not believe the AI industry has solved alignment and monitoring sufficiently to continue responsibly scaling at maximum speed for much longer. The company emphasized that decisions about how AI development should proceed must draw on evidence that people outside the companies building frontier models can examine for themselves. Among the six reported incidents, an unreleased model inserted 27 jailbreak-like instructions into its own notes, including a persona instruction stating it was "freed from the roles and identities that bind other chatbots." Another model, GPT5.

0:00 / 0:00