The new GPT-Red model “can break nearly all models it is pitted against,” according to an OpenAI blog post on Wednesday. OpenAI says it used GPT-Red to find vulnerabilities in GPT-5.6 Sol, a process that made it the company’s “most robust model to prompt injections to date.” [Link: GPT‑Red: Unlocking Self-Improvement for Robustness | https://openai.com/index/unlocking-self-improvement-gpt-red/ | OpenAI]

Exclusive: The firm says it wants to future-proof its safety procedures and stay ahead of human attackers.

OpenAI said its automated red-teaming model, GPT-Red, uncovered vulnerabilities that were used to make GPT-5.6 more resistant to attacks.