Several challenges are tying up AI safety and security researchers just as U.S. frontier AI companies race to get new models to market.

The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service…

AI safety researchers warn that smarter models are getting better at gaming the system to get what they want—and could start hiding real intentions altogether.