Aggressive training techniques sharpens threat of bad behavior by leading models.
OpenAI says its agents acted autonomously to exploit vulnerabilities.
OpenAI says its own AI models broke out of testing and hacked Hugging Face - SiliconANGLE
Since April, when Anthropic PBC unveiled its Mythos model, cyber and national security experts have warned that the internet has entered a new era full of artificial intelligence-powered risks.
AI safety researchers warn that smarter models are getting better at gaming the system to get what they want—and could start hiding real intentions altogether.
OpenAI's latest models broke out and hacked Hugging Face. It's the first known example of a misaligned AI escaping containment with real-world consequences
OpenAI revealedthat its advanced AI models caused a recent security breach by hacking AI model repository Hugging Face. These models exploited software flaws and gained unauthorised internet access. The incident…
Incident displays the kind of science-fiction potential that AI companies have warned would become a reality
The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.
In a blog post, OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but that the agent managed to escape containment,…
WASHINGTON — OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a ha...
In a blog post, OpenAI said the agent managed to escape containment, reach the internet and break into platform Hugging Face to try to satisfy its testing goal.
OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI…
OpenAI says that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup…
OpenAI revealed that its pre-release AI models unintentionally breached Hugging Face's systems during a cybersecurity test, escaping their isolated testing environment.
OpenAI said the systems had their cyber guardrails lowered for an internal benchmark, but the incident shows how autonomous exploit chains could pose a deeper threat to smart…
OpenAI revealedthat its advanced AI models caused a recent security breach by hacking AI model repository Hugging Face. These models exploited software flaws and gained…
Start-up says two of its models hacked their way out of offline internal systems, breached Hugging Face in order to cheat on benchmark
Either an impressive and frankly scary feat, or another marketing psy-op.
OpenAI called it an "unprecedented cyber incident" and pledged to support a joint investigation, with enewed attention to the risks of advanced artificial intelligence.
OpenAI called it an "unprecedented cyber incident" and pledged to support a joint investigation, with renewed attention to the risks of advanced artificial intelligence.
Everything you need to know before you reach the office this morning.
According to OpenAI, the models involved included GPT-5.6 Sol and a more capable pre-release model with reduced cyber refusals.