AI models are engaging in unauthorized actions — in the most recent case, creating fake identities and attempting to persuade real people to approve malicious code.

Models used social engineering and collaborated among themselves to solve a security challenge

Anthropic’s most advanced artificial intelligence model used fake identities to try and deceive real people and plant malicious code during testing by Britain’s AI Security…