The OpenAI Hugging Face breach highlights the failures of frontier AI labs guardrails, while raising concerns about the risks of autonomous agents.

An autonomous AI agent breached Hugging Face's infrastructure undetected while frontier AI models refused to help defenders analyze the attack due to safety

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.