At
Transcript

Statement on the US government directive to suspend access to Fable 5 and Mythos 5 \ Anthropic

At 5:21pm ET, the US government ordered Anthropic to kill Fable 5 for every customer on Earth — not because the model broke, but because a single jailbreak demoed a capability already available from G

@AnthropicAI · anthropic.com

Gist

1.

At 5:21pm ET, the US government ordered Anthropic to kill Fable 5 for every customer on Earth — not because the model broke, but because a single jailbreak demoed a capability already available from GPT-5.5. If one narrow vulnerability can trigger a global shutdown, the frontier AI industry just learned that its deployment window is measured in hours, not months.

Logic

2.

The order came without explanation and hit everyone

  • The government cited "national security authorities" but provided no specific details of its concern — Anthropic's understanding is that a jailbreak was demonstrated
  • The directive covers every foreign national, inside or outside the US, including Anthropic's own employees — a scope that forces a full shutdown to ensure compliance
  • Access to all other Anthropic models is unaffected, isolating the action to Fable 5 and Mythos 5 alone

3.

The jailbreak is narrow, non-universal, and already public

  • Anthropic reviewed the demonstrated technique and found it identifies "a small number of previously known, minor vulnerabilities" — capabilities already available from GPT-5.5 and used daily by defenders
  • No tester has found a universal jailbreak — one that broadly bypasses safeguards across a wide range of cyber tasks
  • The government's verbal evidence describes asking the model to read a codebase and fix flaws — a task Anthropic validated as widely available from other models

4.

Anthropic built defense in depth because perfect resistance is impossible

  • Thousands of hours of red-teaming with the US government, UK AISI, and private organizations showed Fable's safeguards are "substantially more effective than those of any previously deployed model"
  • Anthropic adopted a defense-in-depth strategy: make non-universal jailbreaks narrow and universal ones expensive, then monitor and shut down attacks
  • 30-day data retention was imposed at real customer cost specifically to research and mitigate jailbreaks — a policy no competitor requires

5.

The standard applied here would halt the entire industry

  • Anthropic argues a narrow, non-universal jailbreak should not trigger a global recall of a model deployed to hundreds of millions of people
  • If this standard holds, every frontier model provider faces the same risk — one minor finding, one verbal directive, one 5:21pm order, and the product dies
  • Anthropic publicly supports government authority to block unsafe deployments, but only through a "transparent, fair, clear, and grounded in technical facts" statutory process — this action meets none of those criteria

Counter-Argument

6.

Anthropic is grading its own homework — and the teacher just failed it

  • The entire defense rests on Anthropic's self-assessment: thousands of hours of red-teaming, safeguards "so strong many users have complained they are overly broad," no harmful results disclosed to them. The government's directive is the first independent, external judgment on Fable's safety — and it is a negative one.
  • Anthropic's own launch blog post conceded that "universal jailbreaks will eventually be found in the future" and that "perfect jailbreak resistance does not appear to be possible today." The government may be acting on the timeline Anthropic itself predicted, not on a new standard.
  • The "halts the industry" argument is a bargaining position, not a principle. Anthropic wants a statutory process — but a statutory process is slow, public, and gives the industry time to lobby. The government's 5:21pm directive is fast, opaque, and gives the industry no time to fight. The real dispute is over speed and transparency, not over whether the government should act at all.

Steelman

7.

The directive is not a safety order — it is a sovereignty claim

  • Both Anthropic and its critics share a hidden assumption: that the government's action is about Fable's safety. But the directive covers foreign nationals inside the United States, cites export control authorities, and arrived with no technical explanation — these are the hallmarks of a sovereignty assertion, not a safety finding.
  • The US government is signaling that it can reach inside any company's product, anywhere on Earth, and disable access for anyone it deems a national security risk — without disclosing the threat, without a statutory process, and without a technical standard. This is not a model-specific order; it is a precedent-setting claim over the global AI supply chain.
  • The real question is not whether Fable 5 is safe enough. It is whether any frontier model provider can operate under a regime where a single government can issue a global kill order at 5:21pm based on undisclosed evidence. Anthropic's customers in London, Tokyo, and Berlin are not being protected from a jailbreak — they are being reminded that their access to the most powerful AI models depends on the goodwill of one nation's export control authorities.

Original

Continue Reading