Anthropic's text watermarking alters word probabilities to leave fingerprints, which could degrade Claude's writing, despite its claim of no impact on quality
When I wrote this week about Anthropic's announcement that all Claude models, worldwide, would soon begin “watermarking” …
Daring Fireball John Gruber
Context & Ripple Effects
Anthropic has moved from announcing worldwide Claude watermarking and file-level C2PA metadata for EU AI Act compliance to explaining that the text signal only indicates Claude was likely involved. Its watermark design is sparse in code and factual text and can be removed by a full rewrite, limiting it as a durable authorship check.
The new criticism puts the mechanism itself under scrutiny: changing token probabilities to create a fingerprint creates a potential quality trade-off just months after users challenged Claude performance changes that Anthropic denied were capacity-driven.
First-order effects
- Claude outputs subject to the watermark will be generated from altered word probabilities, making Anthropic's no-quality-impact claim a concrete point of evaluation for Claude users.
- Anthropic gains a provenance signal across Claude text, but the signal is not a definitive attribution tool and is weakest in code and factual material.
Second-order effects
- Teams using Claude in document workflows have an incentive to revise outputs when quality matters; a full rewrite also removes the watermark, reducing the marker's usefulness after ordinary editing.
- Buyers seeking evidence of AI involvement will need to treat Claude's text watermark alongside file metadata rather than as a standalone detector, because the text signal does not establish authorship.
Third-order effects
- If probability-based watermarking becomes a standard compliance route, model providers will face a recurring product-design trade-off between output quality and provenance signals that survive real-world editing.
- The pattern favors provenance systems that disclose likely model involvement over claims of permanent AI-text detection, since editable text can shed the marker.
The trend: Generative-AI compliance is shifting toward layered provenance signals, while the durability and quality cost of those signals become central product questions.
Related: Provenance Over Detection · The Synthetic-Media Trust Stack · Anthropic · Anthropic details Claude's text watermark · Claude users accuse Anthropic of degrading performance
Related Coverage
- Anthropic's Claude to watermark AI-generated text Information Age · Tom Williams
- Claude Users Cancel Subscriptions Over Anthropic's Invisible Text Watermark Implicator.ai · Marcus Schuler
- Anthropic reveals how Claude secretly watermarks AI-written text Android Authority · Hadlee Simons
- Claude's Watermarks and their Legal Sector Impact Artificial Lawyer
- Anthropic's weak watermarks appease a weak law James Padolsey's Blog · James Padolsey
- AI Watermarks and the Trust Layer VC Cafe · Eze Vidra
- Anthropic's New AI Watermark Sparks Backlash From Claude Subscribers Inc.com · Kevin Haynes
- Anthropic Explains Its Watermark System as Some Claude Users Loudly Revolt Gizmodo · Mike Pearl
- Lies And Scams Taint Watermark Removal Apps Now That Anthropic Started Watermarking Claude AI Outputs Forbes · Lance Eliot
- Claude is getting ambitious with watermarking, and I can smell the problems from a mile away Digital Trends · Sudhanshu Kumar Mangalam
- Anthropic is watermarking text generated by Claude to comply with EU law Engadget · Mariella Moon
- Hey, Your AI Is Showing New York Magazine · John Herrman
- Anthropic details Claude's text watermark: it only shows Claude was likely involved, is sparse in code and factual text, and disappears after a full rewrite Anthropic
- Anthropic Reveals What The Watermark Is And How It Can Be Defeated Search Engine Journal · Roger Montti
- Anthropic adds AI watermarks Hard Coded · Emil Protalinski
- Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing Hacker News
- “This isn't just about text one might generate with the intention of passing it off as their own natural work. This isn't even about LLM proofreading of work written by hand. Anthropic is saying that all new Claude models are going to adulterate every single bit of text longer than 200 tokens (~150 words) they generate, including everything it presents to its users to read. … @remixtures@tldr.nettime.org · Miguel Afonso Caetano
- Anthropic explains how Claude's invisible text watermarks will work The Verge · Jess Weatherbed
- Anthropic watermarks Claude's output, but critics question the tradeoffs The Decoder · Maximilian Schreiner
- AI-generated text should be detectable, but Apple needs to avoid Anthropic's huge error 9to5Mac · Ben Lovejoy
- Anthropic's ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing Lobsters
- Why LinkedIn doesn't really care about AI slop MarTech · Constantine von Hoffman
- Claude AI Faces Wave Of Cancellations Over New Invisible Watermark Feature Ubergizmo · Paulo Montenegro
- LLMs aren't writing — I'm not an AI hater: I've used it for my fair share of projects … Six Colors · Dan Moren
- Claude text watermarks will “nudge” its word choices. Should we care? PCWorld · Ben Patterson
- You Gotta Fight For Your Write Spyglass · M.G. Siegler
- What Claude's AI text watermark actually does Mashable · Chance Townsend
- Anthropic explains how Claude's controversial new ‘watermarking’ feature works The Independent · Andrew Griffin
- Claude is changing how it generates prose to be more detectable (and maybe worse?) Nieman Lab · Joshua Benton
- John Gruber calls Claude's AI watermarking ‘patently offensive’ Business Insider · Ben Shimkus
- Anthropic to Add Text Watermarks for Claude's AI Output WinBuzzer · Markus Kasanmascheff
Discussion
-
@arturovilla
Arturo Villa
on x
🔴 So Anthropic just wrote a statement just to confirm all concerns about watermarking. - It's permanent and not removable. - It will be applied to all new models from now on, and existing after December. - They will apply it also to translations. - They will provide a binary AI
-
@seltaa_
@seltaa_
on x
This is exactly why Anthropic pisses me off. You can keep saying the watermark doesn't change ownership, doesn't identify the user, doesn't affect quality, doesn't add hidden characters, and won't look any different to readers. Fine. But that is not the whole issue. The issue is
-
@kimmonismus
@kimmonismus
on x
From now on, if you use Claude just to translate text into another language, it will be considered AI-generated. This has serious consequences, for example, for students at universities, because detectors will now flag it as AI content.
-
@bleeding_sun
Lia
on x
If AI is so amazing, why do they want so bad to hide using it? 🙂 I'll tell you: when you rely on a machine to think so you can pretend you're smart, of course you're insecure about it. Because you can't actually do anything yourself. And now everyone will know.
-
@trq212
@trq212
on x
Watermarking without quality loss is a bit unintuitive, doesn't feel like it should work. I made this artifact with Claude to help me understand how it works, sharing in case it's helpful.
-
@wesroth
Wes Roth
on x
Future Claude models are getting an invisible text watermark. But not the way you'd think. No hidden characters. No extra tokens. No visible label. No user-identifying information. It's basically a statistical pattern hidden in the model's word choices that Anthropic can check
-
@elder_plinius
@elder_plinius
on x
sounds like Anthropic's watermarking is MUCH more interesting than hiding invisible Unicode, as most people assumed. the watermark is baked into the token selection itself! at a high level: normal Claude: context → probability cloud + random sampling → token watermarked
-
@basedjensen
@basedjensen
on x
Anthropic bro's just can't help shoot themselves in the foot with bad comms
-
@nickadobos
Nick Dobos
on x
Anthropic's response to watermarks confirmed my worst fears. “Our watermarking method doesn't have any practical impact on the quality or content” That is a lie. They literally say they change the wording in the first paragraph Anyone with an ounce of common sense about how
-
@kimmonismus
@kimmonismus
on x
Seriously, that post made it worse not better.
-
@jessuppi
Jesse Nickles
on x
The fucking E.U. Every single time, it's fucking Germany destroying the open internet. If you're a blogger, web developer, or digital marketer and you don't hate the German government then you're either uniformed or masochistic...
-
@_mattata
Remy
on x
Hey claude, reply with 1000 “A's” and nothing else Yeah. Think about what you're being told critically please.
-
@hammer_mt
Mike Taylor
on x
Strikes me as fundamentally hostile to writers. There's a huge difference between choosing the word “grey” vs “overcast”!
-
@gnukeith
Keith
on x
“Watermarks can't be traced to a specific person, organization, or chat.” It will, in the future. It always happens.
-
@alex_verem
Alex Veremeyenko
on x
“Let's steal all the data from the internet and then add a watermark to it”
-
@dedene
Peter Dedene
on x
Scrape the entire open web, then watermark the remix. 💀 Another win for local open models.
-
@lyraintheflesh
Lyra Intheflesh
on x
Ahhh... So Claude now uses loaded dice. Got it. Still terrible, even with your explanation. Here's the thing you need to understand: it is ok and good for people to reject this. They can understand the mechanism perfectly and still say, “No. This is not something that is OK.”
-
@rasbt
Sebastian Raschka
on x
A short illustration of how the Claude's watermarking is supposed to work (based on my read of their released materials). In general, when we are generating tokens, there can be multiple high-scoring tokens at certain next-word positions. Usually, we sample with top-k or top-p
-
@jrobertlennon.com
J. Robert Lennon
on bluesky
Claude itself is already a perversion of writing—my writing, that I've written and published over thirty years, that it stole from me—and if it is to be allowed to exist at all, its loser users should forever be forced to bear the mark of shame. What it makes isn't “writing” at …
-
@zzypt
AJ Thomas
on bluesky
I really think your anti-EU hysteria is blinding you to how people feel about AI. People want to know when it's been used, they want to avoid it. These methods seem weak to me but they are wanted.
-
@nymag.com
@nymag.com
on bluesky
Hey, your AI is showing. AI detectors, watermarks, and slop-reporting buttons are coming. John Herrman explores what happens when AI can no longer hide.
-
@frankpasquale
Frank Pasquale
on bluesky
“TikTok, YouTube, and even Meta have taken steps to label some AI-generated content, while LinkedIn, where many users haven't encountered a human word in months, introduced a “Seems Like AI” button. — nymag.com/intelligence...
-
@carnage4life
Dare Obasanjo
on bluesky
John Gruber dislikes Anthropic's watermarking because Claude may choose words that make AI text easier to detect instead of simply writing the best answer. — I personally don't care. Claude's writing was already too clunky to send without editing. If editing also defeats the …
-
@adamfishercox.com
Adam Fisher-Cox
on bluesky
If you're using AI to adjust your writing in the first place, precise word choice and ownership of that isn't important to you, no? Who is the person who is both affected by and cares about this?
-
@gruber.foo
John Gruber
on bluesky
It's unacceptable for a tool to sacrifice an iota of clarity, coherence, meaning, quality, etc. for the purpose of embedding hidden clues within the text to suggest its provenance. — daringfireball.net/2026/08/anthro...
-
@hern
Alex Hern
on bluesky
I'm convinced, upon rereading, that this simply misunderstands how either or both of how LLMs or watermarking works daringfireball.net/2026/08/anth...
-
@metacurity.com
Cynthia Brumfield
on bluesky
This puts the death knell on Claude's usage by writers of all stripes for all purposes, imho. — Anthropic's ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing — daringfireball.net/2026/08/anth...
-
@binaryape
Pete Birkinshaw
on bluesky
This is like complaining about the quality of sesame seeds on your turd sandwich. *AI* is a perversion of writing.
-
@dbreunig
Drew Breunig
on x
This is a good take. I'm not against AI watermarking. I am against the absurd notion that word choice doesn't matter. The confidence with which they make this claim partially explains why Claude has become such a terrible writer. https://daringfireball.net/...
-
r/technology
r
on reddit
Anthropic's ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing
-
r/ArtificialInteligence
r
on reddit
Anthropic's ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing
-
@davidcrespo
@davidcrespo
on bluesky
this strikes me as completely nuts. it's like the computer person version of AI spouse types losing their shit when 4o was being shut down, except they were actually losing something, whereas Gruber is fully inventing the thing he is mad about
-
@matthew.flux.community
Matthew Sheffield
on bluesky
It's hard to imagine why someone would be angry at AI text watermarking. This technology is essential to combating AI slop and also to the companies themselves to prevent model collapse. — And yet, here is this knee-jerking anti-EU piece: daringfireball.net/2026/08/anth...
-
r/LocalLLaMA
r
on reddit
Daring Fireball: Anthropic's ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing
-
@michae.lv
Michael Veale
on bluesky
yes, this blog is v wrong, a shame because he's said some sensible & useful things in the past especially regarding apple app store policies & politics. ai watermarking interestingly invented by visiting researcher at OpenAI many years ago; co built it then sat on it as scared i…
-
@tbertran
@tbertran
on bluesky
I usually agree with his takes but this one is weird to me. Are you seriously getting mad at a machine for writing ever so slightly more mechanical prose? — If authenticity is that important why use AI in the first place.