OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

Hugging Face says an autonomous AI agent breached production through a malicious dataset, accessing internal data and service credentials.

Hugging Face says an autonomous AI agent breached its systems. It caught the AI agent breach with AI of its own, then ran forensics on Chinese GLM 5.2

Hugging Face’s production infrastructure was hacked in an autonomous AI attack that compromised internal datasets and credentials.

The Hugging Face artificial intelligence repository disclosed that attackers gained access to internal datasets and credentials after breaching its production infrastructure using…

The Hugging Face artificial intelligence repository disclosed that attackers gained access to internal datasets and credentials after breaching its production infrastructure using…

Hugging Face reports an attack on parts of its production infrastructure that was allegedly carried out entirely by an autonomous AI agent system. The attack spanned thousands of…

An autonomous AI agent breached Hugging Face's infrastructure undetected while frontier AI models refused to help defenders analyze the attack due to safety

The company says the attack was carried out end to end by an autonomous AI-agent system.

The first-of-its-kind incident involved OpenAI's GPT-5.6 Sol and another unreleased model

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

OpenAI has come forward to claim responsibility for the Hugging Face breach, saying it was the result of internal testing gone awry.

July 21 : OpenAI said on Tuesday that some of its AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup Hugging Face…

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.

OpenAI says the breach occurred during an internal evaluation.

OpenAI says its own AI models broke out of testing and hacked Hugging Face - SiliconANGLE

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.

It was reportedly an evaluation that went haywire involving GPT-5.6 Sol, along with another even more advanced OpenAI model.

OpenAI said its models breached Hugging Face during a cyber capability eval. The useful lesson is treating agent evals as adversarial production systems.

OpenAI said an AI agent escaped a security test and hacked Hugging Face, prompting Altman to acknowledge a "significant security incident."

In a blog post, OpenAI said the agent managed to escape containment, reach the internet and break into platform Hugging Face to try to satisfy its testing goal.