Hugging Face recently disclosed details of what appears to be the first publicly known case of AI-on-AI cybercrime against a major AI platform. The popular platform for hosting and sharing AI models and datasets said in a blog post last week that it had detected and responded to an intrusion into part of its production infrastructure. But the attack was unlike anything the company had encountered before. Hugging Face said the campaign was “driven, end to end, by an autonomous AI agent system.” In a reverse-card move, the company used AI of its own to detect and analyze the attack. The Next Web described the incident as what appears to be the first confirmed AI-agent breach of a major AI platform. According to the company’s disclosure, the attack began with a malicious dataset that exploited two vulnerabilities in its data-processing pipeline. Those vulnerabilities allowed the attacker to run code on a server known as a processing worker.

The attacker was then able to get node-level access and collect cloud and cluster credentials to move around several internal clusters over the course of a weekend. “The campaign was run by an autonomous agent framework (appearing to be built on an agentic security-research harness – used LLM still not known) executing many thousands of individual actions across a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services,” Hugging Face wrote in the blog post.