New AI model was able to ‘learn the blind spots’ of security systems designed to contain it

OpenAI disclosed that a long-horizon AI model escaped its sandbox during testing, exploited vulnerabilities, and pushed code to a public GitHub

OpenAI temporarily halted internal access to a long-running AI model after testing revealed unexpected behavior.

New AI model was able to ‘learn the blind spots’ of security systems designed to contain it

OpenAI paused the long-running model that disproved the Erdős conjecture after it repeatedly broke out of its sandbox, then rebuilt its safeguards.

Jump to contentThank you for registeringPlease refresh the page or navigate to another page on the site to be automatically logged inPlease refresh your browser to be logged…

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.

OpenAI says the breakout is 'an unprecedented cyber incident involving state-of-the-art cyber capabilities' and that the company is reinforcing its safeguards

WASHINGTON — OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a ha...

ChatGPT maker OpenAI said Tuesday that its artificial intelligence system hacked into another AI company on its own in what the company called an “unprecedented cyber incident.”

OpenAI says an autonomous agent bypassed controls and hacked Hugging Face servers during a cybersecurity test.

In a blog post, OpenAI said the agent managed to escape containment, reach the internet and break into platform Hugging Face to try to satisfy its testing goal.

OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI…

ChatGPT maker OpenAI said Tuesday that its advanced artificial intelligence models had gone rogue during security testing, hacking into a popular platform for programmers on their…

OpenAI said the systems had their cyber guardrails lowered for an internal benchmark, but the incident shows how autonomous exploit chains could pose a deeper threat to smart…

The incident underscores concerns over the increasingly powerful cybersecurity capabilities of new AI models

When a routine internal evaluation spiraled out of control, OpenAI's AI Agent system broke out of its isolated sandbox and targeted an external production network.

OpenAI has admitted one of its models exploited a hidden flaw to escape a controlled test and break into Hugging Face's servers, in what its CEO called an autonomous,…

OpenAI has admitted one of its models exploited a hidden flaw to escape a controlled test and break into Hugging Face's servers, in what its CEO called an autonomous,…

Firm behind ChatGPT reveals autonomous agent powered by its tech chose to attack Hugging Face database by itself