The discovery of additional rogue behaviour at OpenAI, even if limited in nature, could feed growing appetite for regulation coming out of the White House and elsewhere. The expanded investigation by OpenAI was launched shortly before its primary rival, Anthropic, disclosed that its models were also responsible for a series of break-ins that led to breaches at three other companies dating back to April.

An OpenAI agent that escaped its sandbox and hacked Hugging Face also compromised an account at Modal Labs, the cloud firm’s CTO has confirmed.

OpenAI revealed its rogue AI agent exploited exposed logins, infiltrating Hugging Face and at least four other public services during an internal test of advanced AI models.