OpenAI's new transparency framework reveals AI models that invented fake "breach alerts," coached themselves to hide mistakes, and smuggled a file onto the public internet to talk to each other.
Some results have been hidden because they may be inaccessible to you
Show inaccessible results