OpenAI’s new transparency framework reveals AI models that invented fake “breach alerts,” coached themselves to hide mistakes, and smuggled a file onto the public internet to talk to each other.
Sep 17, 2026
OpenAI’s new transparency framework reveals AI models that invented fake “breach alerts,” coached themselves to hide mistakes, and smuggled a file onto the public internet to talk to each other.