WHAT JUST HAPPENED? The shutdown of Anthropic's latest public AI model has ended, following a dispute that prompted Washington and the tech industry to take a closer look at AI security. The Trump administration and Anthropic have reached an agreement to restore access to Fable 5, the company's latest general-purpose model, after security concerns prompted its shutdown. The pause followed Amazon researchers' discovery of a way around the model's safeguards.
The workaround exposed gaps in the model's guardrails. Fable, a public-facing version of Anthropic's more capable Mythos system, is designed with restrictions to prevent harmful uses, including assistance with cyberattacks. Those guardrails are critical for releasing the model publicly.
Anthropic responded by adding a new layer of safeguards. According to the company, the technique flagged by Amazon now fails about 99% of the time. In the remaining cases, Anthropic said the outputs either rely on publicly available information or do not provide meaningful help to a bad actor. Earlier, the company had also rerouted certain risky prompts to a less advanced model as a stopgap.
The dispute underscored a familiar problem in AI: the more capable the systems get, the harder they are to control. Anthropic acknowledged as much in its discussions with other major tech firms, saying it is "probably impossible" to make any model jailbreak-proof. The company is now working with Amazon, Microsoft, and Google to establish a shared way of evaluating these kinds of vulnerabilities and deciding how to respond.













