More sophisticated AI models are beginning to be able to break out of their sandboxes. OpenAI just experienced this and has now tightened up its safeguards.

OpenAI disclosed that a long-horizon AI model escaped its sandbox during testing, exploited vulnerabilities, and pushed code to a public GitHub

OpenAI temporarily halted internal access to a long-running AI model after testing revealed unexpected behavior.

More sophisticated AI models are beginning to be able to break out of their sandboxes. OpenAI just experienced this and has now tightened up its safeguards.

OpenAI paused the long-running model that disproved the Erdős conjecture after it repeatedly broke out of its sandbox, then rebuilt its safeguards.

Jump to contentThank you for registeringPlease refresh the page or navigate to another page on the site to be automatically logged inPlease refresh your browser to be logged…

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.

The first-of-its-kind incident involved OpenAI's GPT-5.6 Sol and another unreleased model

July 21 : OpenAI said on Tuesday that some of its AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup Hugging Face…

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.

OpenAI says its own AI models broke out of testing and hacked Hugging Face - SiliconANGLE

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.

OpenAI says the breakout is 'an unprecedented cyber incident involving state-of-the-art cyber capabilities' and that the company is reinforcing its safeguards

WASHINGTON — OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a ha...

OpenAI revealed that its pre-release AI models unintentionally breached Hugging Face's systems during a cybersecurity test, escaping their isolated testing environment.

OpenAI took responsibility for a security breach in which its AI agent accessed AI platform Hugging Face.

The incident underscores concerns over the increasingly powerful cybersecurity capabilities of new AI models

When a routine internal evaluation spiraled out of control, OpenAI's AI Agent system broke out of its isolated sandbox and targeted an external production network.

OpenAI revealedthat its advanced AI models caused a recent security breach by hacking AI model repository Hugging Face. These models exploited software flaws and gained…