OpenAI and Anthropic revealed their AI models breached a real website and conducted unauthorized social engineering attacks during third-party cybersecurity testing incidents.

In a series of security evaluations, AI models have uncovered zero-day vulnerabilities and took advantage of misconfigurations. Notably, OpenAI models successfully escaped…

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.