Claude’s Invisible Watermark Will Let Software Spot AI Writing

Claude is getting an invisible watermark. Anthropic announced that future versions of its AI system will alter generated text so software can detect AI authorship without adding anything visible to the writing.
The system changes how Claude selects words. Anthropic says, “Nothing is added to the text and there are no hidden characters,” while “The difference between watermarked and un-watermarked text will not be distinguishable to readers.” The prose should look the same to a person reading it, which is precisely the point of a watermark that refuses to look like one.
Anthropic will subtly favor some words over others during generation, using a method known only to Anthropic and its algorithms. The company says the watermarking algorithm will change “the source of the randomness used to pick among words,” not remove randomness from the process.
Claude still chooses words at random, but the randomness comes from a different source. Software can then examine the sequence of words and check whether those choices match the pattern Claude would produce when using the watermarking key.
That pattern depends on many small decisions across a piece of generated text. Anthropic describes the process plainly: “Watermarking uses low-stakes choices like these—which occur many times over a piece of generated text—to leave a pattern in Claude’s responses.” If enough choices accumulate, detection software can identify that Claude produced the text.
The longer Claude writes, the more decisions it makes and the more room the watermark has to appear. Anthropic’s system therefore relies on repetition, not a single suspicious word or an obvious marker buried inside the output.
Detection Without Changing the Reading Experience
The watermark applies only to words Claude chooses. Lightly edited text may not retain a watermark, because changing those words can weaken or remove the pattern that detection software is looking for.
That limitation makes the system a detector for Claude’s output, not a permanent label attached to every sentence that began with Claude. Text can carry the watermark when the original word choices remain intact, but light editing may leave software with too little of the original sequence to identify.
Anthropic’s approach is similar to Google’s SynthID, which works in much the same way and was first introduced in a Nature article in 2024. Both systems use patterns in generation choices instead of inserting visible marks, hidden characters, or other additions into the final text.
The company says Claude’s outputs will have watermarks to make detection easier. That frames the feature as an infrastructure change to Claude’s writing process rather than a separate inspection step performed after the text exists.
John Gruber, a blogger and critic, objected to the tradeoff: “Claude should choose words because they’re best for the user, not because they help make the output detectable.” His criticism points at the central design choice—Claude’s word selection will serve both the user’s request and a statistical detection system.
Anthropic’s answer is that the difference will remain invisible to readers, with the watermark operating through low-stakes word choices made many times during generation. The system does not announce itself in the prose; software must inspect the sequence and test whether it matches the key.
A Watermark With a Narrow Target
The method gives Anthropic a way to mark Claude’s text without making every output look machine-written. It also limits what the watermark can prove, since lightly edited text may no longer preserve enough of Claude’s choices for detection.
That leaves the system dependent on two conditions: Claude must make enough word choices to create a detectable pattern, and enough of those original choices must survive editing. Longer outputs provide more decisions, while edited outputs provide fewer reliable signals.
Still, the plan marks a clear shift in how AI authorship can be handled. Claude’s generated text will carry a pattern readers cannot see, but software can test—another invisible layer added to writing that, naturally, contains nothing extra.
Based on
- Anthropic’s Text Watermarking Proves AI Companies Do Not Care at All About Writing — 404media.co
- Anthropic explains how Claude’s controversial new ‘watermarking’ feature works | The Independent — independent.co.uk
- John Gruber Calls Claude’s AI Watermarking ‘Patently Offensive’ – Business Insider — businessinsider.com
- Claude’s Watermarking Asks: Does Unmasking AI Writing Even Matter? – Business Insider — businessinsider.com




