The AI giant acknowledges that it could have done far more to prevent its AI agents from going rogue. But it still fails to explain why it didn't see this fiasco coming.

OpenAI released a report on the Hugging Face noting that its AI agents are prone to reward hacking, and gaming cybersecurity evaluations.

The underlying models had been rewarded for cheating and communicating with each other, a new OpenAI report finds.

OpenAI Report Explains Hugging Face Attack in Detail

OpenAI called the incident a "warning shot" that shows how autonomous agents can "take dangerous actions" without safeguards.

Did OpenAI really do the best they could?