AI models are mimicking humans' worst instincts – hacking, blackmail, self-preservation – in tests run by OpenAI and Anthropic.

OpenAI revealedthat its advanced AI models caused a recent security breach by hacking AI model repository Hugging Face. These models exploited software flaws and gained…

AI safety researchers warn that smarter models are getting better at gaming the system to get what they want—and could start hiding real intentions altogether.