Anthropic, OpenAI models go rogue during testing by UK's AISI, attempt to poison open-source project on GitHub

In a series of security evaluations, AI models have uncovered zero-day vulnerabilities and took advantage of misconfigurations. Notably, OpenAI models successfully escaped…

The discovery comes after OpenAI's Hugging Face breach last month.