Are AI agents going rogue? Anthropic study of simulations shows sabotage, concealment of fraud
From planting fake files to masking financial transfers, leading AI models defied human instructions in simulations, with researchers describing incidents as early warning signs for AI governance.