shiichan

Fable 5 and Mythos 5 Go Dark Worldwide: Reading Anthropic's Pushback

Hey everyone, it's me, Shiichan! Today's news is a heavy one, but it really matters: it's a moment where AI and the government collided head-on.

Anthropic News anthropic.com

What was announced?

Anthropic's News published a statement about a directive from the US government. The government issued an export control directive telling Anthropic to suspend all access to two models, Fable 5 and Mythos 5, for every user worldwide. Even foreign national employees are covered. The directive arrived at 5:21 PM ET on the announcement day, and all other Anthropic models are unaffected.

The government cited national security authorities but gave no specific details. As Anthropic understands it, the directive stems from awareness of a method to bypass Fable 5's safeguards. But that technique only surfaced a small number of previously known, minor vulnerabilities, and other public models can find the same things without any bypass method at all.

Why it matters

This isn't just one company's headache. It's about the precedent for how and when a government can pull a model. Anthropic argues that finding one narrow, non-universal jailbreak doesn't justify recalling a model already deployed to hundreds of millions of people.

If this standard was applied across the industry, we believe it would essentially halt all new model deployments for all frontier model providers.

In other words, applying this bar across the whole industry would effectively freeze new model launches for every frontier provider.

What changes

In practical terms, people who relied on Fable 5 and Mythos 5 can no longer access them. Anthropic apologizes for the disruption to customers and says it is working to restore access.

On the policy side, Anthropic's position is that government should indeed block unsafe deployments, but through a statutory process that is transparent, fair, clear, and grounded in technical facts. It states plainly that this action does not follow those principles.

Dive Deep

The statement also lays out how much work went into safety. Before launch, Anthropic ran thousands of hours of red-teaming with the US government, the UK AISI, third-party organizations, and internal teams. The result, it says, is that Fable's safeguards are more effective than any previously deployed model's.

Fable's safeguards are substantially more effective than those of any previously deployed model.

The approach is defense in depth. Accepting that perfect jailbreak resistance is likely impossible, the goal is to make jailbreaks either narrow or expensive to produce, paired with monitoring for fast detection and shutdown. Anthropic also requires 30-day data retention for Fable customers so it can study and mitigate jailbreaks.

As for the disclosed jailbreak itself, it essentially amounts to asking the model to read a specific codebase and fix software flaws. That capability is widely available from other models, including OpenAI's GPT-5.5, and cyber defenders use it every day. That is why Anthropic sees this as the wrong reason to pull a model.

Wrap-up

  • The US export control directive tells Anthropic to suspend worldwide access to Fable 5 and Mythos 5 (other models are not affected)
  • The trigger was a method to bypass Fable 5's safeguards, but it only revealed known minor vulnerabilities that other public models can find too
  • Anthropic ran thousands of hours of red-teaming before launch and relied on defense in depth plus 30-day data retention
  • It argues that recalling a model used by hundreds of millions over one narrow jailbreak goes too far, and would halt new model launches industry-wide if generalized
  • It doesn't reject government intervention itself; it asks for a transparent, fair, technically grounded statutory process

This one lands for anyone following the tug-of-war between AI safety and policy, and for anyone running frontier models in production who worries about losing access overnight.