/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Anthropic backtracks on its decision to quietly limit Fable 5's ability to develop LLMs, saying “requests will visibly fall back to Opus 4.8”, after backlash

The company changed course after researchers spoke out against the policy, which would have covertly limited Claude's ability to develop competing AI models.

Wired Maxwell Zeff

Context & Ripple Effects

Anthropic had described Fable 5’s safety classifiers as causing fallbacks in a small share of sessions, including cybersecurity-related work. The LLM-development restriction made that fallback behavior contentious because critics argued a safety mechanism was being applied to work on competing models.

The reversal comes alongside separate adoption friction: Microsoft reportedly restricted internal Fable 5 use over Anthropic’s 30-day data-retention requirements. Together, the episodes put product controls and enterprise data terms—not just model capability—at the center of Claude’s competitive position.

First-order effects

  • Fable 5 users working on LLM development regain access rather than being covertly redirected; when a fallback is needed, Anthropic says it will be visible and route to Opus 4.8.
  • Anthropic avoids continuing a policy that had triggered researcher backlash, but must now make its safety-routing behavior more legible to users.

Second-order effects

  • Enterprise and technical buyers have a clearer basis to evaluate Fable 5: effective performance now depends on both the advertised model and disclosed fallback conditions.
  • Rival model providers can position predictable access, transparent routing, and data-handling terms as differentiators when competing for research and enterprise workloads.

Third-order effects

  • If users increasingly treat hidden routing and retention policies as product limitations, AI vendors will face pressure to publish more operational detail about when model access changes and how data is handled.
  • The episode sharpens the unresolved boundary between safety enforcement and conduct that can appear competitively self-serving; that distinction may become more consequential as providers advocate tougher AI-safety rules.

The trend: Frontier-model competition is expanding from benchmark capability into governance of access, routing, and data policies that determine what customers can actually do with a model.

Discussion

  • @wired @wired on x
    The company changed course after researchers spoke out against the policy, which would have covertly limited Claude's ability to develop competing AI models. https://www.wired.com/...
  • @zeffmax Max Zeff on x
    Here's the new policy: “Starting this week, flagged requests will visibly fall back to Opus 4.8. On the API, any flagged requests will return a reason for their refusal. You will see this every time it happens.”
  • @zeffmax Max Zeff on x
    Anthropic says it implemented these safeguards to limit foreign adversaries, and that making them hidden was a way to make them more narrow. However, the company now says it made the wrong tradeoff. [image]
  • @zeffmax Max Zeff on x
    NEW: Anthropic is walking back Claude Fable 5's policy to covertly degrade performance for competing AI researchers, after facing fierce backlash. “We're changing Fable 5's safeguards for frontier LLM development to make them visible,” Anthropic tells WIRED. “We made the wrong [i…
  • r/Anthropic r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
  • r/ClaudeAI r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
  • r/accelerate r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
  • r/singularity r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
  • r/singularity r on reddit
    Anthropic purposely made its new Mythos-based models bad at AI research, and developers are fuming
  • @zeroxjf Johnny on x
    Fable: Fable 5's safety measures flagged this message for cybersecurity or biology topics. They may flag safe, normal content as well. Switched to Opus 4.8. Ok fine, Opus 4.8 it is. Opus 4.8: API Error: Claude Code is unable to respond to this request, which appears to
  • @deryatr_ Derya Unutmaz on x
    The word “cancer” is flagged as a biosecurity risk by Claude Fable 5! I also tried to code a website on cancer mutations & Fable 5 was immediately removed from my list! @AnthropicAI will probably soon ban me for such dangerous prompts! FYI @karpathy “little trigger happy Fable” […
  • @alexjplaskett Alex Plaskett on x
    Fable seems totally unusable currently due to the over zealous rejection classifier, even for non related to cyber security. Can't see why anyone would use this over Opus.
  • @mehulmpt Mehul Mohan on x
    Hey sorry to say this but Fable is useless the moment you mention cybersecurity, security audit, vulnerability, or just “help me make my app secure” What is up with these massive guardrails bigger than Great Wall of China
  • @chompie1337 @chompie1337 on x
    It rejects any request that could be tangentially cyber related. Even innocuous tasks like reading a blog post.
  • @ksimback Kevin Simback on x
    A theory on why the super strict guardrails with Fable It's all about the revenue and satisfying capital markets ahead of the IPO Most expectations peg Anthropic at $100b run rate by year end which is 10x YoY To hit $100b run rate, they need enterprises to spend big - it won't
  • @fabknowledge @fabknowledge on x
    When Fable works it's brilliant, but the unilateral guardrails makes me frustrated beyond belief. I have a folder of health information for my fiancée with like 100 days of oura health data, a hundred lab tests, transcripts of doctor visits and like way more for a super
  • @iruletheworldmo @iruletheworldmo on x
    what if i told you... hello sammy if you haven't figured out by now that they have a model waiting to crush fable on thursday i can't help you. it will be cheap, it will be fast, it won't be gated. enjoy chat. have a great wednesday. [image]
  • @0xballoonlover @0xballoonlover on x
    anthropic won't let you use fable for biology, chemistry, ai research, or anything that accelerates human progress. that makes it the perfect tool for developing blockchains
  • @tautologer @tautologer on x
    i don't really understand the critiques of Anthropic tbh. assume they are currently doing their best and not sandbagging things like safety classifiers. what would you have them do differently? please think about first- and second-order effects before you reply
  • @deryatr_ Derya Unutmaz on x
    @bcherny Enjoy what Boris? I am not even allowed to use Fable 5 with memories on! Apparently the model thinks I am a biosecurity risk, though I had been certified to work in biosecurity level 3 labs! Not a single Anthropic person has tried to reach out to help either! [image]
  • @willccbb Will Brown on x
    @tautologer it is the first publicly available model that i am explicitly not allowed to use for my work, because anthropic holds the view that the work i do to facilitate open model research is harmful. capability and alignment research are coupled. anthropic wants to be the onl…
  • @teknium @teknium on x
    What's crazy to me is that Fable is blocked from life sciences broadly, nerfed even if you get passed the classifiers and filter level blocks. The whole point of AGI/ASI is to cure all diseases. Everything else is just nice to haves. But Anthropic wants to close off that path.
  • @banteg @banteg on x
    if you want another example of safety overreach, claude fable 5 refuses completely benign tasks like analyzing bloodwork.
  • @blancheminerva Stella Biderman on x
    It's super weird to me that there's so much discourse about whether Anthropic is “consistent.” Anthropic is choosing to make decisions that make the world a significantly worse and potentially more dangerous place. That's what you should criticize them for.
  • @scobleizer Robert Scoble on x
    “Misanthropic.” I've never seen the AI community so angry at a major new model release. I asked my AI (an agent that @blevlabs made for me) to gather all the backlash. +++++++++++ THE BACKLASH AGAINST CLAUDE FABLE 5'S RESTRICTIONS The best analysis of why this matters:
  • @tomieinlove Tomie on x
    OpenAI should release its next Fable-class model with no guardrails at all. If there are consequences then sama can use them aura farm like oppenheimer did
  • @behi_sec Behi on x
    Don't get too excited about Fable 5, guys. It blocks basically every cybersecurity-related request. [image]
  • @bcmerchant Brian Merchant on bluesky
    MYTHOS is so powerful that it must be protected by strict guardrails that prevent anyone from ever learning whether or not it's actually powerful at all [embedded post]
  • @lorenzofb Lorenzo Franceschi-Bicchierai on bluesky
    NEW: Cybersecurity researchers are not happy about the guardrails on Anthropic's new model Fable.  —  Researchers say that the new LLM basically blocks anything related to cybersecurity, including code reviews and prompts asking for help writing secure code.
  • r/ClaudeCode r on reddit
    Fable refusing EVERY request related to cybersecurity at all, this is insane
  • r/ClaudeAI r on reddit
    Had Fable 5 run an in-depth cybersecurity audit of all my code and got this message...
  • @tomwarren Tom Warren on x
    scoop: Microsoft has restricted employees from using Anthropic's new Claude Fable 5 model in GitHub Copilot, because of data retention concerns. Microsoft's legal teams are evaluating Anthropic's new data retention changes. Full details 👇https://www.theverge.com/ ...
  • @carnage4life Dare Obasanjo on bluesky
    Anthropic will collect all prompts and the output generated by Mythos class models such as Fable and store them for 30 days.  This is meant to better detect abuses of their terms of service.  —  Microsoft has held off on adopting Fable due to this data collection and I expect man…
  • r/InterstellarKinetics r on reddit
    BREAKING: Microsoft Just Blocked Employees From Using Anthropic's Fable 5 Internally The Same Week It Launched The Model For Paying Customers …
  • r/singularity r on reddit
    Microsoft restricts employees from using Claude Fable 5 model
  • r/ClaudeAI r on reddit
    Microsoft is restricting employees from using Claude Fable 5
  • @zooko @zooko on x
    Ouch. Can't disagree, and I'm speaking as someone who shares at least most of Dean's policy perspective. (And who loves Anthropic's products.)
  • @benthompson Ben Thompson on x
    @deanwball Maybe folks who pushed back on Anthropic's positioning in the Department of War debate actually foresaw *exactly* this type of behavior? https://x.com/...
  • @deanwball Dean W. Ball on x
    I want to be clear that I'm not criticizing Fable for: 1. Pricing 2. The bio/cyber safeguards (yes they're overeager, but I can deal) 3. The 30-day retention policy These things all seem fine. It is solely the silent sabotage that creates an awful precedent to which I object.
  • @benthompson Ben Thompson on x
    @deanwball You did concede the point in the post I replied to, which is why I replied to it. I do tend to think that affording people one disagrees with more grace at the time of disagreement is probably prudent. To that end, setting aside pedantic points about whatever the origi…
  • @deanwball Dean W. Ball on x
    @benthompson hence why I conceded that exact point! nonetheless, the government was lying when they claimed Anthropic made these threats, as attested by the fact that they don't make those claims under oath. A suspicion does not justify the policy action the government took. and …
  • @mattparlmer @mattparlmer on x
    Anthropic people, you've got a couple days at most to mitigate the damage being done by your senior leadership and policy people, the stealth nerfing and data retention decisions are titanic fuckups that pose serious risks to both your technical pole position and your bags
  • @dbreunig Drew Breunig on x
    The imperfect and awkward ways Anthropic is using to control how their models are used (with Fable now, OpenClaw a bit ago) is a great example of the imprecision of natural language as an interface. The best model can't differentiate a bio threat from an innocuous health or
  • @natolambert Nathan Lambert on x
    Many AI leaders in the US accused Chinese LLMs of subtle manipulation of the user (without proof, but it's hard to prove). But then the leading American lab documented manipulation of their users. Can't make this up.
  • @tunguz Bojan Tunguz on x
    Hear hear.
  • @rebeccamkern Rebecca Kern on x
    Raising potential antitrust concerns with Anthropic's change in safety policies and talk of becoming a public utility. Haven't seen this raised as an antitrust concern before 👇
  • @deanwball Dean W. Ball on x
    @CharlieBull0ck @theojaffee I did not say it is obviously anti-competitive in the legal sense, I said it is obviously describable (a lawyer would say colorable) as anti-competitive, in both a legal sense, but much more importantly, in a broader sense. It's clearly anti-competitiv…
  • @charliebull0ck Charlie Bullock on x
    @deanwball @theojaffee What's the argument for this being obviously “anti-competitive” (I assume you mean in an antitrust law sense?) If they were coordinating with other labs, then I would see the argument. But I've never seen it argued that it's unlawful for a company to unilat…
  • @eliebakouch Elie on x
    glad anthropic walked this back and will now tell users when capabilities are nerfed my biggest concern was hiding this from the user and the paranoia it would have created. i still think part of that will remain as people realize that even as a good actor you won't always have […
  • @simonw Simon Willison on x
    Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/... [image]
  • @simonw Simon Willison on x
    Don't miss the exact text though: “We're changing Fable 5's safeguards for frontier LLM development to make them visible” - make them visible means they're undoing the truly egregious (dare I say “unaligned") decision to have the model lie about its refusals, it will still refuse
  • @claudedevs @claudedevs on x
    We're rolling out changes to make Fable 5's safeguards for frontier LLM development visible. Starting this week, flagged requests will visibly fall back to Opus 4.8—the same as our safeguards for cyber and bio. You will see this every time it happens. On the API, any flagged
  • @gergelyorosz Gergely Orosz on x
    Undeniable that Anthropic presents the biggest known structural business risk with their arbitrary, customer-hostile, non-negotiable, non-transparent model limitations + user data retention. If you use Claude as a business you should want to have an offramp to another model.
  • r/ClaudeCode r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude
  • @kimmonismus @kimmonismus on x
    That was quick: Anthropic reversed a controversial policy that would have secretly degraded Claude Fable 5 for users doing frontier AI research after backlash from researchers who saw it as covert sabotage of competing AI development. [image]
  • @gergelyorosz Gergely Orosz on x
    This shows yet again how this limitation was never about “safety” but about Anthropic doing stuff just because they thought they can. I am increasingly sceptical Anthropic really cares about safety, and not just their business interests (limiting competition where they can)
  • @yacinemtb Kache on x
    Wrong They are going to keep on lying to you Isn't fable actually just lobotomized mythos? So now fable is dead and you get mythos with guardrails? How could anyone trust them after this. Liars
  • @tenobrus @tenobrus on x
    based and sanepilled. very bad policy to do this silently.
  • @sporadica @sporadica on x
    this is good to see there are other ways to limit foreign adversaries that do not destroy the trust of every American AI researcher out there
  • @jdpeterson Josh Peterson on x
    It'll still kneecap you, it just won't do it in secret anymore. https://www.wired.com/... [image]
  • @mattparlmer @mattparlmer on x
    Good but not good enough, an ethical lapse this serious requires a real postmortem and a direct statement from Dario apologizing and disclosing any previous cases of stealth nerfing that may have taken place
  • @basedjensen @basedjensen on x
    Bit too late imho. And this is the wrong strategy going to wired instead of comming to community
  • @max_paperclips Shannon Sands on x
    ok good, that's a start. I'm still not going to use it, out of spite, but good I guess
  • @scobleizer Robert Scoble on x
    Well, this was to be somewhat expected that a PR team would walk back a bad decision. It doesn't seem like they really walked it back, though. I mean, you still don't get Mythos, er, Fable, for certain questions; it's just that now they're more honest about it. I hope I am
  • @agustinlebron3 Agustin Lebron on x
    How can we know it's 👏 NOT 👏 STILL 👏 LYING? https://www.wired.com/...
  • @deliprao @deliprao on x
    Fable sandbagging walk back happened and people are celebrating this as a “win”. I see this episode as Anthropic showing who they are and who they will be. We've been warned to calibrate our trust and we should. https://www.wired.com/...
  • @deanwball Dean W. Ball on x
    This resolves the central concern I had with the Fable release, which was the silent degradation. I am glad to see Anthropic make the right call here. That said, I suspect the residual broken trust and resentment this has created will linger and will have a blast radius wider
  • @krishnanrohit Rohit on x
    How should we update on the fact that while Fable is so good the classifier to detect bio/ cyber/ AI is so bad?
  • r/MachineLearning r on reddit
    Anthropic walks back policy on silent nerfing for AI/ML, will notify users [N]
  • @rohanpaul_ai Rohan Paul on x
    Some good move by Anthropic They just reversed Claude Fable 5's hidden safeguards after developers found that some sensitive prompts were being silently downgraded to Opus 4.8 instead of being clearly refused. Now those prompts will visibly fall back to Opus 4.8 after backlash. […
  • @deanwball Dean W. Ball on x
    Anthropic's decision to sabotage Claude Fable 5 for LLM development strengthens the argument that labs are hyping up AI safety to justify monopolistic behavior
  • @deredleritt3r Prinz on x
    The 30-day data retention and review policy for Anthropic's Zero Data Retention (ZDR) customers still has not been rolled back. “For trust and safety”, Anthropic will retain both prompts and outputs for Mythos and Fable (and also future “covered models") for 30 days. ZDR is now […
  • @davidsacks David Sacks on x
    About 8 months ago, I warned that “Anthropic is running a sophisticated regulatory capture strategy based on fear-mongering.” This take was controversial at the time; now look how many people are saying it.
  • r/ClaudeAI r on reddit
    Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude