Anthropic backtracks on its decision to quietly limit Fable 5's ability to develop LLMs, saying “requests will visibly fall back to Opus 4.8”, after backlash
The company changed course after researchers spoke out against the policy, which would have covertly limited Claude's ability to develop competing AI models.
Wired Maxwell Zeff
Context & Ripple Effects
Anthropic had described Fable 5’s safety classifiers as causing fallbacks in a small share of sessions, including cybersecurity-related work. The LLM-development restriction made that fallback behavior contentious because critics argued a safety mechanism was being applied to work on competing models.
The reversal comes alongside separate adoption friction: Microsoft reportedly restricted internal Fable 5 use over Anthropic’s 30-day data-retention requirements. Together, the episodes put product controls and enterprise data terms—not just model capability—at the center of Claude’s competitive position.
First-order effects
- Fable 5 users working on LLM development regain access rather than being covertly redirected; when a fallback is needed, Anthropic says it will be visible and route to Opus 4.8.
- Anthropic avoids continuing a policy that had triggered researcher backlash, but must now make its safety-routing behavior more legible to users.
Second-order effects
- Enterprise and technical buyers have a clearer basis to evaluate Fable 5: effective performance now depends on both the advertised model and disclosed fallback conditions.
- Rival model providers can position predictable access, transparent routing, and data-handling terms as differentiators when competing for research and enterprise workloads.
Third-order effects
- If users increasingly treat hidden routing and retention policies as product limitations, AI vendors will face pressure to publish more operational detail about when model access changes and how data is handled.
- The episode sharpens the unresolved boundary between safety enforcement and conduct that can appear competitively self-serving; that distinction may become more consequential as providers advocate tougher AI-safety rules.
The trend: Frontier-model competition is expanding from benchmark capability into governance of access, routing, and data policies that determine what customers can actually do with a model.
Related: Anthropic · Claude · Anthropic says Claude Fable 5 uses conservative safety classifiers tha · Sources: Microsoft is limiting Claude Fable 5 use internally due to An
Related Coverage
- Why Claude switched models in your conversation with Fable 5 Claude
- Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming Business Insider · Alistair Barr
- Anthropic's Fable a cautionary tale Runtime · Tom Krazit
- After backlash, Anthropic to revise Claude Fable 5's AI development restrictions Moneycontrol · Vikas SN
- Anthropic debuts Claude Fable 5 with gated access, risking a martech gap eMarketer · Grace Harmon
- Claude Mythos 5 Can Build Exploits But Can't Power Campaigns BankInfoSecurity.com · Michael Novinson
- Anthropic's Mythos goes public—on a short leash Tech Brew · Lindsey Choo
- Anthropic accused of ‘secret sabotage’ as Claude Fable 5 silently limits capabilities for AI researchers and developers Fortune · Sharon Goldman
- Fable 5: Guardrails and burn rate are annoying users, who say it's still better than Opus 4.8 The New Stack · Meredith Shubel
- Anthropic's Claude Fable 5 is out for public use, with safeguards for high-risk requests Help Net Security · Sinisa Markovic
- Claude Fable won't answer basic biology questions The Verge · Robert Hart
- Fable 5 and the Shift From Model Capability to Capability Governance Chirag Mehta June 10, 2026 … Constellation Research · Chirag Mehta
- Anthropic's New Fable AI Model Is Met With User Backlash Over Restrictions Wall Street Journal · Sam Schechner
- Scientists angered as Anthropic AI shuts them out Telegraph · Matthew Field
- Anthropic's Mythos Safeguards Stoke Fears of a ‘Permanent Underclass’ Gizmodo · Webb Wright
- Anthropic finally releases Mythos to the public, but it's so heavily guarded it barely works How-To Geek · Jon Fingas
- NEW: Cybersecurity researchers are not happy about the guardrails on Anthropic's new model Fable. — Researchers say that the new LLM basically blocks anything related to cybersecurity, including code reviews and prompts asking for help writing secure code. — “[Fable] rejects any request that could be tangentially cyber related. … @lorenzofb@infosec … · Lorenzo Franceschi-Bicchierai
- Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable Hacker News
- Announcements — Claude Fable 5 introduces our 5th model generation for your most ambitious work. Anthropic
- Microsoft limits employee use of Anthropic's Claude Fable 5 over data retention concerns, The Verge reports Reuters · Juby Babu
- Claude Fable 5 is generally available for GitHub Copilot The GitHub Blog
- Microsoft is restricting Claude Fable 5 for staff despite its AI ambitions, here is why Digit · Ayushi Jain
- How Fable 5 And Mythos 5 Change AI Security, Data Retention, And Vendor Risk Forrester · Jeff Pollard
- Anthropic Unveils Mythos 5 Upgrade Exclusive for Tech Industry and ‘Safe’ Version for the Public International Business Times · Matias Civita
- MSFT Restricts Internal Use Of Claude Fable Over Data-Retention Concerns; BMO Calls Anthropic A Leading Pure-Play AI Lab ZeroHedge News · Tyler Durden
- Microsoft: OpenAI Exposure Manageable And Palantir Competition Contained Seeking Alpha · Bashar Issa
- Microsoft Balks at Anthropic's Claude Fable 5 Data Retention Policy PYMNTS
- Anthropic opens powerful Claude Fable 5-AI model to general users The American Bazaar · Kashmira Konduparty
- Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs The Decoder · Matthias Bastian
- OpenAI and Anthropic renew AI safety warnings amid race for growth Türkiye Today
- Real-time cyber safeguards on Claude Claude
- Anthropic: ‘We made the wrong tradeoff’ in new model guardrails Business Insider · Shubhangi Goel
- Claude Fable 5: Anthropic admits “wrong tradeoff” after invisibly throttling rival AI researchers The Decoder · Maximilian Schreiner
- Anthropic is walking back its plan to silently degrade Claude Fable 5 if it detects you're trying to use it to build a competing LLM. — Instead Claude will either refuse the request or reroute the user to a less capable model. — Imagine if Microsoft had done this with Office or Visual Studio. … @carnage4life@mas.to · Dare Obasanjo
- Vibe Coding With Claude Fable 5 BridgeMind on YouTube
- Anthropic backtracks on policy that ‘sabotaged’ researchers' work Engadget · Steve Dent
- Anthropic Apologizes For One of the Guardrails on Its Fable 5 Model, and Will Change It Gizmodo · Mike Pearl
- If Claude Fable stops helping you, you'll never know (via) Jonathon Ready highlights … Simon Willison's Weblog · Simon Willison
- ‘Oh God, no! Not another thing:’ What Anthropic's Mythos-class Fable 5 means for CEOs trying to govern AI Fortune · Diane Brady
- Anthropic apologizes for invisible Claude Fable guardrails The Verge · Robert Hart
- Fable 5 backlash: Company silently makes this important change for researchers and developers Moneycontrol · MC Learn
- Anthropic rolls out ‘Mythos-like’ AI model Claude Fable 5 Silicon Republic · Laura Varley
- Anthropic Reveals Full Guide to Prompting Claude Fable 5: Why Traditional AI Prompting No Longer Works WinCentral · Nisha
- Microsoft warns employees against using Anthropic Claude Fable 5: Says 'Legal is still evaluating Anthropic's...' Moneycontrol · MC Learn
- Sources: Microsoft is limiting Claude Fable 5 use internally due to Anthropic's 30-day data retention rules, due to customer data and confidential info concerns The Verge · Tom Warren
- Cybersecurity researchers complain that Claude Fable 5's guardrails are too restrictive, rejecting “innocuous tasks” like reading blog posts or reviewing code TechCrunch · Lorenzo Franceschi-Bicchierai
- Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude (via) Big scoop for Maxwell Zeff at Wired: … Simon Willison's Weblog · Simon Willison
- The Moral Of Fable — On June 9, Anthropic released Claude Fable 5, the first of its Mythos … Forbes · Christian Catalini
- Anthropic Was So Concerned About Its New Mythos-Based Model's Power That It Lobotomized Its Ability to Improve Itself Futurism · Victor Tangermann
- Microsoft limits employee use of Claude Fable 5 over data retention concerns TechRadar · Craig Hale
- Claude Fable 5 is Anthropic's most capable public AI model, and will hand your conversation to a weaker model the moment it detects a biology or chemistry question … Silicon Canals
- Anthropic Makes Claude Fable Guardrails Visible After Apology WinBuzzer · Markus Kasanmascheff
- Anthropic Apologizes for Claude Fable 5 Secret Censorship—But the Fix Has a Catch Decrypt · Jose Antonio Lanz
- Anthropic's Claude Fable 5 plays it too safe on safety, developers say Fast Company · Mark Sullivan
- Anthropic rankles users with safety-first Fable release NBC News · Jared Perlo
Discussion
-
@wired
@wired
on x
The company changed course after researchers spoke out against the policy, which would have covertly limited Claude's ability to develop competing AI models. https://www.wired.com/...
-
@zeffmax
Max Zeff
on x
Here's the new policy: “Starting this week, flagged requests will visibly fall back to Opus 4.8. On the API, any flagged requests will return a reason for their refusal. You will see this every time it happens.”
-
@zeffmax
Max Zeff
on x
Anthropic says it implemented these safeguards to limit foreign adversaries, and that making them hidden was a way to make them more narrow. However, the company now says it made the wrong tradeoff. [image]
-
@zeffmax
Max Zeff
on x
NEW: Anthropic is walking back Claude Fable 5's policy to covertly degrade performance for competing AI researchers, after facing fierce backlash. “We're changing Fable 5's safeguards for frontier LLM development to make them visible,” Anthropic tells WIRED. “We made the wrong [i…
-
r/Anthropic
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
-
r/ClaudeAI
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
-
r/accelerate
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
-
r/singularity
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
-
r/singularity
r
on reddit
Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming
-
@zeroxjf
Johnny
on x
Fable: Fable 5's safety measures flagged this message for cybersecurity or biology topics. They may flag safe, normal content as well. Switched to Opus 4.8. Ok fine, Opus 4.8 it is. Opus 4.8: API Error: Claude Code is unable to respond to this request, which appears to
-
@deryatr_
Derya Unutmaz
on x
The word “cancer” is flagged as a biosecurity risk by Claude Fable 5! I also tried to code a website on cancer mutations & Fable 5 was immediately removed from my list! @AnthropicAI will probably soon ban me for such dangerous prompts! FYI @karpathy “little trigger happy Fable” […
-
@alexjplaskett
Alex Plaskett
on x
Fable seems totally unusable currently due to the over zealous rejection classifier, even for non related to cyber security. Can't see why anyone would use this over Opus.
-
@mehulmpt
Mehul Mohan
on x
Hey sorry to say this but Fable is useless the moment you mention cybersecurity, security audit, vulnerability, or just “help me make my app secure” What is up with these massive guardrails bigger than Great Wall of China
-
@chompie1337
@chompie1337
on x
It rejects any request that could be tangentially cyber related. Even innocuous tasks like reading a blog post.
-
@ksimback
Kevin Simback
on x
A theory on why the super strict guardrails with Fable It's all about the revenue and satisfying capital markets ahead of the IPO Most expectations peg Anthropic at $100b run rate by year end which is 10x YoY To hit $100b run rate, they need enterprises to spend big - it won't
-
@fabknowledge
@fabknowledge
on x
When Fable works it's brilliant, but the unilateral guardrails makes me frustrated beyond belief. I have a folder of health information for my fiancée with like 100 days of oura health data, a hundred lab tests, transcripts of doctor visits and like way more for a super
-
@iruletheworldmo
@iruletheworldmo
on x
what if i told you... hello sammy if you haven't figured out by now that they have a model waiting to crush fable on thursday i can't help you. it will be cheap, it will be fast, it won't be gated. enjoy chat. have a great wednesday. [image]
-
@0xballoonlover
@0xballoonlover
on x
anthropic won't let you use fable for biology, chemistry, ai research, or anything that accelerates human progress. that makes it the perfect tool for developing blockchains
-
@tautologer
@tautologer
on x
i don't really understand the critiques of Anthropic tbh. assume they are currently doing their best and not sandbagging things like safety classifiers. what would you have them do differently? please think about first- and second-order effects before you reply
-
@deryatr_
Derya Unutmaz
on x
@bcherny Enjoy what Boris? I am not even allowed to use Fable 5 with memories on! Apparently the model thinks I am a biosecurity risk, though I had been certified to work in biosecurity level 3 labs! Not a single Anthropic person has tried to reach out to help either! [image]
-
@willccbb
Will Brown
on x
@tautologer it is the first publicly available model that i am explicitly not allowed to use for my work, because anthropic holds the view that the work i do to facilitate open model research is harmful. capability and alignment research are coupled. anthropic wants to be the onl…
-
@teknium
@teknium
on x
What's crazy to me is that Fable is blocked from life sciences broadly, nerfed even if you get passed the classifiers and filter level blocks. The whole point of AGI/ASI is to cure all diseases. Everything else is just nice to haves. But Anthropic wants to close off that path.
-
@banteg
@banteg
on x
if you want another example of safety overreach, claude fable 5 refuses completely benign tasks like analyzing bloodwork.
-
@blancheminerva
Stella Biderman
on x
It's super weird to me that there's so much discourse about whether Anthropic is “consistent.” Anthropic is choosing to make decisions that make the world a significantly worse and potentially more dangerous place. That's what you should criticize them for.
-
@scobleizer
Robert Scoble
on x
“Misanthropic.” I've never seen the AI community so angry at a major new model release. I asked my AI (an agent that @blevlabs made for me) to gather all the backlash. +++++++++++ THE BACKLASH AGAINST CLAUDE FABLE 5'S RESTRICTIONS The best analysis of why this matters:
-
@tomieinlove
Tomie
on x
OpenAI should release its next Fable-class model with no guardrails at all. If there are consequences then sama can use them aura farm like oppenheimer did
-
@behi_sec
Behi
on x
Don't get too excited about Fable 5, guys. It blocks basically every cybersecurity-related request. [image]
-
@bcmerchant
Brian Merchant
on bluesky
MYTHOS is so powerful that it must be protected by strict guardrails that prevent anyone from ever learning whether or not it's actually powerful at all [embedded post]
-
@lorenzofb
Lorenzo Franceschi-Bicchierai
on bluesky
NEW: Cybersecurity researchers are not happy about the guardrails on Anthropic's new model Fable. — Researchers say that the new LLM basically blocks anything related to cybersecurity, including code reviews and prompts asking for help writing secure code.
-
r/ClaudeCode
r
on reddit
Fable refusing EVERY request related to cybersecurity at all, this is insane
-
r/ClaudeAI
r
on reddit
Had Fable 5 run an in-depth cybersecurity audit of all my code and got this message...
-
@tomwarren
Tom Warren
on x
scoop: Microsoft has restricted employees from using Anthropic's new Claude Fable 5 model in GitHub Copilot, because of data retention concerns. Microsoft's legal teams are evaluating Anthropic's new data retention changes. Full details 👇https://www.theverge.com/ ...
-
@carnage4life
Dare Obasanjo
on bluesky
Anthropic will collect all prompts and the output generated by Mythos class models such as Fable and store them for 30 days. This is meant to better detect abuses of their terms of service. — Microsoft has held off on adopting Fable due to this data collection and I expect man…
-
r/InterstellarKinetics
r
on reddit
BREAKING: Microsoft Just Blocked Employees From Using Anthropic's Fable 5 Internally The Same Week It Launched The Model For Paying Customers …
-
r/singularity
r
on reddit
Microsoft restricts employees from using Claude Fable 5 model
-
r/ClaudeAI
r
on reddit
Microsoft is restricting employees from using Claude Fable 5
-
@zooko
@zooko
on x
Ouch. Can't disagree, and I'm speaking as someone who shares at least most of Dean's policy perspective. (And who loves Anthropic's products.)
-
@benthompson
Ben Thompson
on x
@deanwball Maybe folks who pushed back on Anthropic's positioning in the Department of War debate actually foresaw *exactly* this type of behavior? https://x.com/...
-
@deanwball
Dean W. Ball
on x
I want to be clear that I'm not criticizing Fable for: 1. Pricing 2. The bio/cyber safeguards (yes they're overeager, but I can deal) 3. The 30-day retention policy These things all seem fine. It is solely the silent sabotage that creates an awful precedent to which I object.
-
@benthompson
Ben Thompson
on x
@deanwball You did concede the point in the post I replied to, which is why I replied to it. I do tend to think that affording people one disagrees with more grace at the time of disagreement is probably prudent. To that end, setting aside pedantic points about whatever the origi…
-
@deanwball
Dean W. Ball
on x
@benthompson hence why I conceded that exact point! nonetheless, the government was lying when they claimed Anthropic made these threats, as attested by the fact that they don't make those claims under oath. A suspicion does not justify the policy action the government took. and …
-
@mattparlmer
@mattparlmer
on x
Anthropic people, you've got a couple days at most to mitigate the damage being done by your senior leadership and policy people, the stealth nerfing and data retention decisions are titanic fuckups that pose serious risks to both your technical pole position and your bags
-
@dbreunig
Drew Breunig
on x
The imperfect and awkward ways Anthropic is using to control how their models are used (with Fable now, OpenClaw a bit ago) is a great example of the imprecision of natural language as an interface. The best model can't differentiate a bio threat from an innocuous health or
-
@natolambert
Nathan Lambert
on x
Many AI leaders in the US accused Chinese LLMs of subtle manipulation of the user (without proof, but it's hard to prove). But then the leading American lab documented manipulation of their users. Can't make this up.
-
@tunguz
Bojan Tunguz
on x
Hear hear.
-
@rebeccamkern
Rebecca Kern
on x
Raising potential antitrust concerns with Anthropic's change in safety policies and talk of becoming a public utility. Haven't seen this raised as an antitrust concern before 👇
-
@deanwball
Dean W. Ball
on x
@CharlieBull0ck @theojaffee I did not say it is obviously anti-competitive in the legal sense, I said it is obviously describable (a lawyer would say colorable) as anti-competitive, in both a legal sense, but much more importantly, in a broader sense. It's clearly anti-competitiv…
-
@charliebull0ck
Charlie Bullock
on x
@deanwball @theojaffee What's the argument for this being obviously “anti-competitive” (I assume you mean in an antitrust law sense?) If they were coordinating with other labs, then I would see the argument. But I've never seen it argued that it's unlawful for a company to unilat…
-
@eliebakouch
Elie
on x
glad anthropic walked this back and will now tell users when capabilities are nerfed my biggest concern was hiding this from the user and the paranoia it would have created. i still think part of that will remain as people realize that even as a good actor you won't always have […
-
@simonw
Simon Willison
on x
Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/... [image]
-
@simonw
Simon Willison
on x
Don't miss the exact text though: “We're changing Fable 5's safeguards for frontier LLM development to make them visible” - make them visible means they're undoing the truly egregious (dare I say “unaligned") decision to have the model lie about its refusals, it will still refuse
-
@claudedevs
@claudedevs
on x
We're rolling out changes to make Fable 5's safeguards for frontier LLM development visible. Starting this week, flagged requests will visibly fall back to Opus 4.8—the same as our safeguards for cyber and bio. You will see this every time it happens. On the API, any flagged
-
@gergelyorosz
Gergely Orosz
on x
Undeniable that Anthropic presents the biggest known structural business risk with their arbitrary, customer-hostile, non-negotiable, non-transparent model limitations + user data retention. If you use Claude as a business you should want to have an offramp to another model.
-
r/ClaudeCode
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
-
@kimmonismus
@kimmonismus
on x
That was quick: Anthropic reversed a controversial policy that would have secretly degraded Claude Fable 5 for users doing frontier AI research after backlash from researchers who saw it as covert sabotage of competing AI development. [image]
-
@gergelyorosz
Gergely Orosz
on x
This shows yet again how this limitation was never about “safety” but about Anthropic doing stuff just because they thought they can. I am increasingly sceptical Anthropic really cares about safety, and not just their business interests (limiting competition where they can)
-
@yacinemtb
Kache
on x
Wrong They are going to keep on lying to you Isn't fable actually just lobotomized mythos? So now fable is dead and you get mythos with guardrails? How could anyone trust them after this. Liars
-
@tenobrus
@tenobrus
on x
based and sanepilled. very bad policy to do this silently.
-
@sporadica
@sporadica
on x
this is good to see there are other ways to limit foreign adversaries that do not destroy the trust of every American AI researcher out there
-
@jdpeterson
Josh Peterson
on x
It'll still kneecap you, it just won't do it in secret anymore. https://www.wired.com/... [image]
-
@mattparlmer
@mattparlmer
on x
Good but not good enough, an ethical lapse this serious requires a real postmortem and a direct statement from Dario apologizing and disclosing any previous cases of stealth nerfing that may have taken place
-
@basedjensen
@basedjensen
on x
Bit too late imho. And this is the wrong strategy going to wired instead of comming to community
-
@max_paperclips
Shannon Sands
on x
ok good, that's a start. I'm still not going to use it, out of spite, but good I guess
-
@scobleizer
Robert Scoble
on x
Well, this was to be somewhat expected that a PR team would walk back a bad decision. It doesn't seem like they really walked it back, though. I mean, you still don't get Mythos, er, Fable, for certain questions; it's just that now they're more honest about it. I hope I am
-
@agustinlebron3
Agustin Lebron
on x
How can we know it's 👏 NOT 👏 STILL 👏 LYING? https://www.wired.com/...
-
@deliprao
@deliprao
on x
Fable sandbagging walk back happened and people are celebrating this as a “win”. I see this episode as Anthropic showing who they are and who they will be. We've been warned to calibrate our trust and we should. https://www.wired.com/...
-
@deanwball
Dean W. Ball
on x
This resolves the central concern I had with the Fable release, which was the silent degradation. I am glad to see Anthropic make the right call here. That said, I suspect the residual broken trust and resentment this has created will linger and will have a blast radius wider
-
@krishnanrohit
Rohit
on x
How should we update on the fact that while Fable is so good the classifier to detect bio/ cyber/ AI is so bad?
-
r/MachineLearning
r
on reddit
Anthropic walks back policy on silent nerfing for AI/ML, will notify users [N]
-
@rohanpaul_ai
Rohan Paul
on x
Some good move by Anthropic They just reversed Claude Fable 5's hidden safeguards after developers found that some sensitive prompts were being silently downgraded to Opus 4.8 instead of being clearly refused. Now those prompts will visibly fall back to Opus 4.8 after backlash. […
-
@deanwball
Dean W. Ball
on x
Anthropic's decision to sabotage Claude Fable 5 for LLM development strengthens the argument that labs are hyping up AI safety to justify monopolistic behavior
-
@deredleritt3r
Prinz
on x
The 30-day data retention and review policy for Anthropic's Zero Data Retention (ZDR) customers still has not been rolled back. “For trust and safety”, Anthropic will retain both prompts and outputs for Mythos and Fable (and also future “covered models") for 30 days. ZDR is now […
-
@davidsacks
David Sacks
on x
About 8 months ago, I warned that “Anthropic is running a sophisticated regulatory capture strategy based on fear-mongering.” This take was controversial at the time; now look how many people are saying it.
-
r/ClaudeAI
r
on reddit
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude