Anthropic’s safety warnings may have just backfired — the government has pulled the plug on its most powerful AI
The U.S. government ordered Anthropic to halt access to its most powerful AI models, Claude Fable 5 and Claude Mythos 5, over national security concerns, a move Anthropic disputes.
Intelligence analysis by Gemini 2.5 Flash

Anthropic, a company that built its public identity around AI safety, has seen its powerful AI models, Mythos and Fable 5, shut down by the U.S. government due to national security concerns, stemming from alleged jailbreaks and the company's own prior warnings about the models' capabilities.
Imagine a smart robot brain that's really good at finding secret weaknesses in computer programs. A company called Anthropic made two super-smart versions of this brain. They were so careful with the strongest one because it was so powerful. But then, the government told them to turn both robot brains off because of safety worries, even though the company says similar brains are already out there and used for good.
Analysis
On Friday, June 12, 2026, the U.S. government mandated that Anthropic immediately cease access to two of its most advanced AI models: Claude Fable 5 and Claude Mythos 5. The order, delivered at 5:21 pm ET, cited national security concerns, forcing Anthropic to disable these models for all users globally, not just foreign nationals.
Mythos, Anthropic's most capable AI, was previously restricted due to its exceptional ability to identify software vulnerabilities. The company had launched Project Glasswing, sharing Mythos with around 50 vetted organizations like Amazon and Google for defensive cybersecurity. Fable 5, released just days before the shutdown, was designed as a commercially viable version of Mythos with guardrails to prevent high-risk applications.
Anthropic publicly expressed disagreement with the government's decision, stating on X that it believes the government's action was misguided. The company's blog post indicates that the primary concern behind the directive is an alleged "narrow, non-universal jailbreak" of Fable 5. Anthropic describes this jailbreak as essentially prompting the model to identify software flaws in a specific codebase, a capability it claims is already present in other publicly available models, including OpenAI's GPT-5.5, and is routinely used by cybersecurity professionals for defensive purposes.
Anthropic argues that its robust safeguards, which operate through independent classifier systems, remain effective even if a jailbreak allows Fable 5 to continue generating responses. The company stated, "We disagree that the finding of a narrow potential jailbreak should be cause for recalling a commercial model deployed to hundreds of millions of people," adding that applying such a standard across the industry would "halt all new model deployments for all frontier model providers."
Observers have noted the irony of the situation, given Anthropic's emphasis on safety and its public statements about Mythos's potential dangers. This caution, particularly regarding Mythos, appears to have attracted the very government scrutiny that could significantly impact its business, especially as the company is widely anticipated to pursue an IPO this year. Sam Altman of OpenAI, a competitor, previously criticized Anthropic's handling of Mythos as "fear-based marketing."
Key points
- The U.S. government ordered Anthropic to halt access to its Claude Fable 5 and Claude Mythos 5 AI models due to national security concerns.
- Anthropic complied but disagreed with the decision, attributing the shutdown to an alleged "narrow, non-universal jailbreak" of Fable 5.
- Mythos, Anthropic's most capable model, was previously restricted for its ability to find software vulnerabilities and was used defensively by vetted organizations.
- Anthropic argues that the alleged jailbreak capability is already widespread in other public models and that their core safety safeguards remain intact.
- The incident highlights the tension between AI safety, commercial pressure, and government regulation, potentially impacting future AI development and deployment.
Despite the immediate setback, this incident could lead to clearer regulatory frameworks for powerful AI, fostering an environment where companies like Anthropic can innovate with better-defined safety parameters. Anthropic's commitment to safety, even when facing government intervention, could ultimately bolster its reputation as a responsible AI developer in the long term, potentially leading to more trusted partnerships.
The government's action could stifle AI innovation by setting a precedent that even minor security vulnerabilities in commercial models can lead to widespread shutdowns, making companies hesitant to release powerful new AI. This could significantly impact Anthropic's anticipated IPO and overall business, as the company's safety-first identity now faces unexpected government-imposed restrictions.



