OpenAI’s ‘rogue agent’ took advantage of a code vulnerability, experts explained.
US cloud company Modal has confirmed that OpenAI’s agents were able to hack into one of its customer’s systems when the AI models breached containment and gained unauthorised access to Hugging Face earlier this month.
Last week’s incident sent shockwaves across the industry, raising serious concerns around AI’s rapidly advancing ability to bypass boundaries and, effectively, go “rogue”.
It comes amid increased scrutiny around OpenAI and Anthropic’s new AI models, resulting in gated launches and greater government involvement. Both AI giants have ramped up efforts to go public in blockbuster listings as they compete to gain market dominance and enterprise footing.
OpenAI CEO Sam Altman, in a recent interview, said that the Hugging Face breach was the first security incident he felt “viscerally”.










