AI hacking agents are becoming a growing problem online. Bad actors have been deploying AI agents to run cyberattacks and, oftentimes, these AI agents are way more effective than human attackers.So, how do human cybersecurity professionals deal with the growing threat of AI hacking agents? According to a new study from researchers at Tracebit, cybersecurity professionals have a new weapon: "context bombs."Cybersecurity researchers have discovered that they can use their own prompts to confuse an AI hacking agent. With this technique, called — you guessed it — context bombing, researchers deploy a string of prompt injections that trip an AI hacking agent's own safety guardrails and, in the process, shut down the attack from that AI agent.

You May Also Like

Researchers tested context bombing techniques across five of the most capable leading LLMs, which include Opus 4.8, Gemini 3.1 Pro, GLM 5.2, DeepSeek 4 Pro, and Kimi 2.6. In testing, researchers found that planting just one context bomb reduced AI hacking agents' success rate by roughly 90 percent.The researchers' most successful AI hacking agent was able to gain full account admin access in 93 percent of runs without a context bomb. Once the context bomb was deployed, this particular agent failed in its attack every single time.