A myriad of software makes up the typical AI harness, and trust issues between the components can create concerning attack vectors.
July 30, 2026
Major frontier AI vendors — including Anthropic, Google, and OpenAI — need to rein in the harnesses they wrap around their large language modules, to limit security weaknesses created by software components that are too trusting of each other.
That's the word from researchers at AI penetration testing firm Novee Security, who were able to use Google's AI agent to execute a supply chain attack and write to its own repository on GitHub, says Elad Meged, a founding team and security researcher at the company. The team also found issues in Anthropic's and OpenAI's AI agents by exploiting misalignments in the trust between elements to enable attacks.
AI harnesses are the software frameworks that provide tools, memory and guardrails for managing AI models; components can include functions like context management, tool integration, and feedback loops too. When a company adopts an AI agent and makes it part of their infrastructure, they are also adopting the trust assumptions of all of those components as well, Meged explains.












