OpenAI says GPT-Red automates prompt injection testing and helped GPT-5.6 Sol record sixfold fewer direct injection failures than GPT-5.5 in benchmark

Exclusive: The firm says it wants to future-proof its safety procedures and stay ahead of human attackers.

OpenAI said its automated red-teaming model, GPT-Red, uncovered vulnerabilities that were used to make GPT-5.6 more resistant to attacks.