L'azienda l'ha chiamato GPT-Red e lo usa come sparring partner per blindare i suoi modelli contro il prompt injection. Contro GPT-5.6 meno di un attacco su quattro va a segno, ma OpenAI ha deciso che non lo render� mai pubblico

Exclusive: The firm says it wants to future-proof its safety procedures and stay ahead of human attackers.

OpenAI said its automated red-teaming model, GPT-Red, uncovered vulnerabilities that were used to make GPT-5.6 more resistant to attacks.