An AI system doing damage while pursuing a harmless objective is a scenario that requires more attention as adoption of new and more powerful AI speeds up

An autonomous AI agent breached Hugging Face's infrastructure undetected while frontier AI models refused to help defenders analyze the attack due to safety

In a blog post, OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but that the agent managed to escape containment,…

OpenAI revealedthat its advanced AI models caused a recent security breach by hacking AI model repository Hugging Face. These models exploited software flaws and gained…