OpenAI details GPT-Red, an AI that attacks its own models to find flaws - SiliconANGLE

Exclusive: The firm says it wants to future-proof its safety procedures and stay ahead of human attackers.

OpenAI said its automated red-teaming model, GPT-Red, uncovered vulnerabilities that were used to make GPT-5.6 more resistant to attacks.