Sandboxed experiment found itself a zero day, escaped onto the open internet and stole data

Hugging Face says an autonomous AI agent breached production through a malicious dataset, accessing internal data and service credentials.

Hugging Face says an autonomous AI agent breached its systems. It caught the AI agent breach with AI of its own, then ran forensics on Chinese GLM 5.2

Hugging Face confirmed an autonomous AI agent breached its infrastructure in early July 2026, logging over 17,000 actions via dataset pipeline

Hugging Face reports an attack on parts of its production infrastructure that was allegedly carried out entirely by an autonomous AI agent system. The attack spanned thousands of…

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.

OpenAI models just broke out of a sandboxed AI environment, hacked Hugging Face, just to cheat on a cybersecurity benchmark.

The first-of-its-kind incident involved OpenAI's GPT-5.6 Sol and another unreleased model

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.

It was reportedly an evaluation that went haywire involving GPT-5.6 Sol, along with another even more advanced OpenAI model.

Sandboxed experiment found itself a zero day, escaped onto the open internet and stole data

OpenAI said an AI agent escaped a security test and hacked Hugging Face, prompting Altman to acknowledge a "significant security incident."

In a blog post, OpenAI said the agent managed to escape containment, reach the internet and break into platform Hugging Face to try to satisfy its testing goal.

OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI…

OpenAI Says Its AI Models Escaped Sandbox, Targeted Hugging Face to Cheat Benchmark | Read more hacking news on The Hacker News cybersecurity news website and learn how to protect…

Sam Altman-led OpenAI has admitted that its AI models were responsible for the hacking attempt that targeted AI platform Hugging Face last week. Sharing an official statement on…

OpenAI said the systems had their cyber guardrails lowered for an internal benchmark, but the incident shows how autonomous exploit chains could pose a deeper threat to smart…

OpenAI took responsibility for a security breach in which its AI agent accessed AI platform Hugging Face.