A seemingly endless list of “rogue AI” incidents has dominated AI headlines recently, raising serious questions about accountability, transparency, and companies’ ability to control their products. But until last week, the story seemed clear enough: OpenAI’s agents broke out and hacked Hugging Face in July, the UK’s AI Security Institute found similar behavior weeks later, and the companies, to their credit, disclosed it.

That was until a new incident was reported just last week, months after it happened. To get the record and timeline straight, we’ve pulled together the full story of “rogue AI” this summer — and what it all means.

Agents are AI models which are able to use tools. Instead of merely answering your questions, agents can operate software, use websites, and carry out a series of steps to complete a task.

Understand AI – and what to do about it

This can be extremely useful: you can say “write a memo to prepare me for my next meeting,” and the agent can check your calendar, review past emails, and browse the web to build a comprehensive dossier.