An AI system doing damage while pursuing a harmless objective is a scenario that requires more attention as adoption of new and more powerful AI speeds up

An autonomous AI agent breached Hugging Face's infrastructure undetected while frontier AI models refused to help defenders analyze the attack due to safety

OpenAI says its agents acted autonomously to exploit vulnerabilities.