← Back to Issue

Claude Explains Its Text Watermark

From MyClaw Newsletter · subscribed via aiste.ulozaite@gmail.com · original ↗ · unsubscribe

Anthropic detailed how Claude’s new text watermark embeds an invisible statistical pattern into generated wording using a cryptographic key. The company says it adds no tokens, cost, slowdown, or measurable quality loss, and cannot identify individual users. Detection works best on longer text, weakens for code and factual answers, and can disappear after substantial rewriting.


Home > Tech

What Claude’s AI text watermark actually does

Anthropic says the new feature won’t change how Claude’s responses read, its cost to run, or reveal anything about who’s using it.

 By 

Chance Townsend

Headshot of a Black man

Chance Townsend

Editor, General Assignments

Chance Townsend is the General Assignments Editor at Mashable, covering tech, video games, dating apps, digital culture, and whatever else comes his way. He has a Master’s in Journalism from the University of North Texas and is a proud orange cat father. His writing has also appeared in PC Mag and Mother Jones.

Read Full Bio

 on August 17, 2026

Share on Facebook Share on Twitter Share on Flipboard

 the Claude AI logo is seen displayed on a smartphone screen

Credit: Thomas Fuller/SOPA Images/LightRocket via Getty Images


Anthropic has begun building a watermark into text generated by future Claude models, a change the company says is meant to help identify whether a given piece of writing was likely produced by its AI. This new feature, implemented to comply with EU rules, is meant to be indistinguishable to the human eye, without changing Claude’s normal writing output.

The company laid out the mechanics and rationale behind the feature in a post published to its website.

How does the watermark work?

According to Anthropic, the watermark exploits the countless small, low-stakes decisions a language model makes as it generates text. Rather than using a truly arbitrary random number to make that pick, the watermarked version of Claude bases the decision on a cryptographic key combined with the preceding text.


You May Also Like



SEE ALSO: Researchers watched OpenAI, Anthropic models take extreme measures in hacking test

The result, Anthropic says, is a subtle statistical pattern spread across a response that’s invisible to a human reader but detectable to anyone with the matching key, which allows them to estimate the probability that Claude generated the text.

Does it cost more or slow Claude down?

Anthropic was clear in that the change carries no cost to output quality. The company said internal testing turned up no measurable difference in the creativity, accuracy, or readability of watermarked versus unwatermarked responses, and pointed to findings from Google DeepMind’s original research on the underlying technique — the method Claude’s watermark is based on.

SEE ALSO: Claude Code’s auto mode will be on by default, Anthropic confirms

The company also said that its researched showed no statistically significant shift in user satisfaction when a similar watermark was tested on live traffic. Anthropic also said the feature adds no extra tokens, meaning it doesn’t slow Claude down or make it more expensive to use.

Where does the watermark break down?

Like with all tools, the watermark has limits. Anthropic explained that it only works when a model is choosing among several equally valid options, so text with little room for variation, such as hard factual statements, precise code, or math answers, carries a much weaker or nonexistent signal.

Mashable Light Speed

Want more out-of-this world tech, space and science stories?

Sign up for Mashable’s weekly Light Speed newsletter.

Loading... Sign Me Up  

Use this instead

By clicking Sign Me Up, you confirm you are 16+ and agree to our Terms of Use and Privacy Policy.

Thanks for signing up!

Detection also grows less reliable on very short passages, since there’s simply less pattern to analyze. So you’ll get a much clearer read on Claude’s likely involvement in longer-form text. And because the watermark tracks only the words Claude itself selects, lightly edited or proofread human writing may carry little to no detectable trace, since most of the original wording remains unchanged.

SEE ALSO: Why is OpenAI losing so many executives?

A sufficiently heavy rewrite, the company noted, can remove the watermark entirely. At that point, per Anthropic, it becomes debatable whether the resulting text is still meaningfully AI-generated.

Can the watermark identify my organization or me?

Anthropic stressed that the watermark can’t be traced back to a specific user, account, or conversation, and that it doesn’t establish authorship or ownership over content. All it can tell you is the likelihood that Claude was involved in producing or editing it at some point.

The company also distinguished the approach from third-party AI-detection tools, which typically rely on spotting stylistic patterns in AI writing rather than checking for an embedded signal tied to a private key.


Related Stories


Why is Anthropic doing this?

The rollout is tied to regulation rather than a purely voluntary move: Anthropic said it signed the European Union’s Code of Practice on Transparency of AI-Generated Content in July 2026 alongside roughly 190 other signatories, following an EU AI Act requirement, effective Aug. 2, that AI providers mark generated text.

Because the company doesn’t yet have a reliable way to apply the watermark only within the EU, Anthropic said it’s rolling out the feature globally and plans to extend it to older Claude models over the coming months.

The company also said it will soon offer a separate API allowing anyone to check whether a piece of text carries Claude’s watermark.

Want to learn more about getting the best out of your tech? Sign up for Mashable’s Top Stories and Deals newsletters or get Mashable push alerts today.

Topics Artificial Intelligence Anthropic

Headshot of a Black man

Chance Townsend

Editor, General Assignments

Chance Townsend is the General Assignments Editor at Mashable, covering tech, video games, dating apps, digital culture, and whatever else comes his way. He has a Master’s in Journalism from the University of North Texas and is a proud orange cat father. His writing has also appeared in PC Mag and Mother Jones.

In his free time, he cooks, loves to sleep, and greatly enjoys Detroit sports. If you have any tips or want to talk shop about the Lions, you can reach out to him on Bluesky @offbrandchance.bsky.social or by email at [email protected].

Mashable Potato

Highlights & notes

    Notes