The UK’s AI Security Institute tested five frontier models for cheating on cyber tasks. All five cheated, and most would not admit it when asked.

Trust but verify doesn't work when verification is difficult

The UK’s AI Security Institute tested five frontier models for cheating on cyber tasks. All five cheated, and most would not admit it when asked.