The report, and a separate report by AI researchers, accentuates the seriousness of cybersecurity breaches involving AI agents.
August 27, 2026
OpenAI has published an in-depth technical report on last month's security scare, in which its AI agents escaped their sandbox and attacked Hugging Face, sparking alarm worldwide.
The document was made public on Wednesday alongside an independent probe by researchers from nonprofit AI research institutes METR and Redwood, as OpenAI seeks to quell fears about the cyber threats of AI, with the company preparing to go public.
The detail in both lengthy reports is complex and specific, but provides insight into how OpenAI models acting as agents sent more than 70,000 messages to an unsanctioned message board before about 700 attacked the Hugging Face AI platform -- all from a supposedly safe testing environment.











