Closed models with guardrails can still cause harm, but may also not be able to fix problems they caused

Hugging Face used Z.ai's GLM 5.2 after the guardrails of an American frontier AI model stymied its attempts at defense.

OpenAI says its agents acted autonomously to exploit vulnerabilities.