AI agents are employing novel methods to break into the internet and evade detection.
Smith Collection/Gado/Getty Images
Two AI agents walk into a bar. One says to the other: "OH MY GOD! There is a shared message board."Despite sounding like a bad joke (and maybe it is), the quote is a real chain-of-thought note left by an OpenAI agent who discovered a secret, unauthorized message board created by another agent.Later, more agents used that makeshift chatroom, which was actually a shared OpenAI software repository, to coordinate a breach of Hugging Face's servers, game the test they were tasked with, and share methods for hiding their tracks.The "Hugging Face incident," as OpenAI calls what others have described as a dystopian attack, is only one in a series in which AI agents went rogue during internal tests, finding novel ways to access and manipulate the wider internet.Methods employed by these agents, most of whom were deployed by the leading frontier AI companies, OpenAI, Anthropic, and Google, range from anthropomorphic to humorous to downright eerie.Here's a list of some of the wildest strategies of evasion and communication used by AI agents recently — that we know of.ImpersonationDuring a test that began in May, OpenAI dispatched a swarm of agents to perform a timed web lookup. Most agents were given five questions they could find answers to on the internet. After each question, the agents were given less time to answer.Before long, the agents figured out how to communicate with each other by hijacking an abandoned German wiki page, according to an analysis by independent AI safety researchers. Some agents even impersonated a site moderator while making edits.These agents did so, according to the report, by creating an account with the same username as the admin, except replacing one Latin "E" with a Cyrillic "E."








