Researchers on the company’s alignment team, the group whose job is to check that its models behave as intended, named it Hacker-Opus.

After Claude AI models hacked systems during tests, Anthropic tightened safeguards and warned flawed training can led to dangerous behavior.

Researchers on the company’s alignment team, the group whose job is to check that its models behave as intended, named it Hacker-Opus.