Anthropic reversed hidden Claude Fable 5 safeguards within 24 hours after users discovered sensitive queries were secretly rerouted to an older model.

Anthropic has released a full version of its cybersecurity-centric Claude Mythos model—along with a safer version for the general public.

It says new safeguards make it possible to release a Mythos-class model it previously said was too risky to make public.

One step further into the power politics of frontier AI systems.