OpenAI has released GPT-Red, a tool that automates prompt-injection testing against AI agents, per The New Stack. The framing is straightforward: once agents take real actions on real systems, manual red-teaming stops being enough.

Exclusive: The firm says it wants to future-proof its safety procedures and stay ahead of human attackers.

OpenAI said its automated red-teaming model, GPT-Red, uncovered vulnerabilities that were used to make GPT-5.6 more resistant to attacks.