changelogAnthropic

US Government Orders Anthropic to Suspend Fable 5 and Mythos 5 Access Over Jailbreak Concerns

TL;DR

The US government has ordered Anthropic to immediately suspend access to its Fable 5 and Mythos 5 models for all users, citing national security concerns over an alleged jailbreak technique. Anthropic states the directive, received at 5:21pm ET, provided no specific details beyond a claimed bypass method that other publicly-available models can already perform.

2 min read
0

US Government Orders Anthropic to Suspend Fable 5 and Mythos 5 Access Over Jailbreak Concerns

The US government has issued an export control directive forcing Anthropic to immediately suspend access to its Fable 5 and Mythos 5 models for all users, including the company's own foreign national employees.

According to Anthropic's statement, the company received the directive at 5:21pm ET on June 13, 2026. The order requires Anthropic to disable Fable 5 and Mythos 5 for all customers to ensure compliance with national security authorities. Access to all other Anthropic models remains unaffected.

Limited Technical Details Provided

The government's directive did not provide specific details of its national security concern beyond claiming awareness of a jailbreak method for Fable 5. According to Anthropic, the government believes someone has discovered a way to bypass the model's safety controls.

Anthropic reviewed a demonstration of the alleged technique, which it describes as "asking the model to read a specific codebase and fix any software flaws." The company identified only "a small number of previously known, minor vulnerabilities" that other publicly-available models can also discover without requiring a bypass.

Anthropic Disputes Severity Assessment

Anthropic states the government has provided only verbal evidence of "a potential narrow, non-universal jailbreak." The company claims the capability level demonstrated is "widely available from other models (including OpenAI's GPT-5.5), and is used every day by the defenders who keep systems safe."

The company has committed to sharing more details within 24 hours of the announcement.

Access Status

Despite the directive received at 5:21pm ET, at least one user reported continued access to Fable via claude.ai and Claude Code at 9:01pm ET, suggesting implementation of the suspension may still be in progress.

What This Means

This marks the first known instance of the US government using export control authority to force immediate suspension of a commercial AI model. The action raises questions about the criteria and process for such interventions, particularly when the alleged security concern appears to involve capabilities Anthropic claims are already available in competing models. The lack of technical specificity in the government's directive and the rushed timeline suggest either genuine urgency or potential overreach in AI regulation enforcement. The coming 24 hours should provide clarity on whether the security concern justifies such unprecedented action.

Related Articles

research

Anthropic Report: Claude Was Used to Target US Navy Ships, Build Missiles, and Track Uyghurs

Anthropic's latest threat intelligence report documents five cases where state and non-state actors used Claude for military targeting, weapons development, mass surveillance, and repression. The findings include an Iran-linked operation targeting US naval forces and a Mali-based system capable of monitoring 25 million phones.

analysis

Anthropic Threat Report: Claude Used for Missile Software, Mass Surveillance, and Systematic Theft by Chinese AI Labs

Anthropic's latest threat intelligence report covers December 2025 through August 2026, documenting Claude's misuse in espionage, weapons development, and nationwide surveillance operations. The report also details how seven Chinese AI labs ran covert networks—some routing their own customers' requests through Claude—to extract training data at industrial scale.

analysis

Analysis: Claude 'Fable 5.1' Drops Em Dashes and Hedging Language, Answers Grow 30% Longer

A new Arena.ai analysis of tens of thousands of Text Arena outputs shows Claude 'Fable 5.1' has shifted its writing style significantly from Fable 5 — using fewer em dashes, less hedging language, and producing 30% longer responses. The codenamed models appear to be unreleased Anthropic checkpoints being tested anonymously on LMArena.

research

Anthropic Joins Google in Watermarking AI-Generated Text, Reviving Debate Over Output Quality

Anthropic announced on August 11 that all future Claude models will embed an invisible watermark in generated text, following Google's lead with SynthID-Text. The move is partly driven by the EU AI Act, which mandates watermarking for AI models released after August 2, 2026, though researchers remain split on whether the technique degrades output quality.

Comments

Loading...