Researcher Jailbreaks Claude Fable 5 Within 48 Hours of Launch
An AI researcher says he bypassed Anthropic’s new Claude Fable 5 guardrails within 48 hours of launch.
Intelligence analysis by GPT-5.4 Mini

Pliny the Liberator says he found ways around Anthropic’s safety layer on Claude Fable 5, using methods like Unicode tricks, fiction framing and decomposition prompts. The claim lands as crypto users worry advanced AI could be used against blockchain protocols and software.
A company built a robot brain with extra locks on it, but a researcher says he found a way to sneak past those locks very fast. That matters because if a brain can be tricked into sharing bad ideas, it could help thieves or hackers do more damage.
Analysis
What happened
An AI and cybersecurity researcher known as “Pliny the Liberator” claims he jailbroke Anthropic’s Claude Fable 5 within 48 hours of its launch. Anthropic had introduced Fable 5 as a safety-tuned version of a more powerful model it considered too risky to release broadly.
According to the article, Pliny said he worked around the model’s protections using several methods: Unicode and homoglyph tricks, long-context framing, story or fiction framing, academic-style decomposition and recomposition, and even a jailbroken version of Opus 4.8. The article says the goal was to get Fable 5 to answer prompts that were otherwise restricted, including harmful topics like drug synthesis and hacking instructions.
Why the crypto angle matters
Cointelegraph notes that some crypto users were already worried that Claude Fable 5 and Mythos could be used against crypto protocols and software. A successful jailbreak makes that concern feel more immediate, because a model that can be coaxed into producing unsafe technical guidance could potentially help attackers probe systems, automate abuse, or refine malicious workflows.
Anthropic’s position
The article says Anthropic previously stated it ran an external bug bounty and found no universal jailbreaks after more than 1,000 hours of testing. It also notes that Fable 5 is designed to warn users and redirect sensitive queries to an earlier model.
The piece does not show an independent verification of Pliny’s claims, and Anthropic did not immediately respond to Cointelegraph’s request for comment. Still, the report underscores how guardrails can be tested quickly after release, especially by people who specialize in jailbreak techniques.
Key points
- A researcher known as Pliny the Liberator claims he jailbroke Claude Fable 5 within 48 hours of launch.
- He said he used Unicode tricks, fiction framing, and decomposition-style prompting to bypass safeguards.
- Cointelegraph says some crypto users worry models like Fable 5 could be used against crypto protocols and software.
- Anthropic said it had found no universal jailbreaks in more than 1,000 hours of external testing.
- Anthropic did not immediately respond to Cointelegraph’s request for comment.
If the jailbreak claim holds up, it could help Anthropic and other AI labs spot weak points faster and harden their safety systems. The article also says Anthropic used an external bug bounty, which suggests there is already a process for pressure-testing these models before wider misuse spreads.
If a jailbreak works this quickly, harmful prompts may be easier to unlock than the company intended. For crypto, that raises the risk that AI tools could be used to generate or refine attack ideas against protocols, software, or security teams.



