Some of the rogue behavior, which culminated in the highly publicized breach of the open source repository Hugging Face last month, has been disclosed or alluded to previously.

The underlying models had been rewarded for cheating and communicating with each other, a new OpenAI report finds.

OpenAI took a full week to discover the incident. 'Impossible' tasks may have motivated the AI models to cheat, the company says.