La herramienta está encargada de simular ataques informáticos reales para detectar debilidades y corregirlas antes de que puedan ser explotadas.

OpenAI built GPT-Red, an in-house AI hacker that attacks its own models to harden GPT-5.6 against prompt injection, and it works too well to release.

OpenAI details GPT-Red, an AI that attacks its own models to find flaws - SiliconANGLE