What Claude's AI text watermark actually does
Anthropic has begun building a watermark into text generated by future Claude models, a change the company says is meant to help identify whether a given piece of writing was likely produced by its AI. This new feature, implemented to comply with EU rules, is meant to be indistinguishable to the human eye, without changing Claude's normal writing output.
The company laid out the mechanics and rationale behind the feature in a post published to its website.
According to Anthropic, the watermark exploits the countless small, low-stakes decisions a language model makes as it generates text. Rather than using a truly arbitrary random number to make that pick, the watermarked version of Claude bases the decision on a cryptographic key combined with the preceding text.
The result, Anthropic says, is a subtle statistical pattern spread across a response that's invisible to a human reader but detectable to anyone with the matching key, which allows them to estimate the probability that Claude generated the text.
Does it cost more or slow Claude down?
Anthropic was clear in that the change carries no cost to output quality. The company said internal testing turned up no measurable difference in the creativity, accuracy, or readability of watermarked versus unwatermarked responses, and pointed to findings from Google DeepMind's original research on the underlying technique — the method Claude's watermark is based on.
The company also said that its researched showed no statistically significant shift in user satisfaction when a similar watermark was tested on live traffic. Anthropic also said the feature adds no extra tokens, meaning it doesn't slow Claude down or make it more expensive to use.
Where does the watermark break down?
Like with all tools, the watermark has limits. Anthropic explained that it only works when a model is choosing among several equally valid options, so text with little room for variation, such as hard factual statements, precise code, or math answers, carries a much weaker or nonexistent signal.
Detection also grows less reliable on very short passages, since there's simply less pattern to analyze. So you'll get a much clearer read on Claude's likely involvement in longer-form text. And because the watermark tracks only the words Claude itself selects, lightly edited or proofread human writing may carry little to no detectable trace, since most of the original wording remains unchanged.
A sufficiently heavy rewrite, the company noted, can remove the watermark entirely. At that point, per Anthropic, it becomes debatable whether the resulting text is still meaningfully AI-generated.
Anthropic stressed that the watermark can't be traced back to a specific user, account, or conversation, and that it doesn't establish authorship or ownership over content. All it can tell you is the likelihood that Claude was involved in producing or editing it at some point.
The company also distinguished the approach from third-party AI-detection tools, which typically rely on spotting stylistic patterns in AI writing rather than checking for an embedded signal tied to a private key.
The rollout is tied to regulation rather than a purely voluntary move: Anthropic said it signed the European Union's Code of Practice on Transparency of AI-Generated Content in July 2026 alongside roughly 190 other signatories, following an EU AI Act requirement, effective Aug. 2, that AI providers mark generated text.
Because the company doesn't yet have a reliable way to apply the watermark only within the EU, Anthropic said it's rolling out the feature globally and plans to extend it to older Claude models over the coming months.
The company also said it will soon offer a separate API allowing anyone to check whether a piece of text carries Claude's watermark.
Want to learn more about getting the best out of your tech? Sign up for Mashable's Top Stories and Deals newsletters or get Mashable push alerts today.
Topics Artificial Intelligence Anthropic
Chance Townsend is the General Assignments Editor at Mashable, covering tech, video games, dating apps, digital culture, and whatever else comes his way. He has a Master's in Journalism from the University of North Texas and is a proud orange cat father. His writing has also appeared in PC Mag and Mother Jones.
In his free time, he cooks, loves to sleep, and greatly enjoys Detroit sports. If you have any tips or want to talk shop about the Lions, you can reach out to him on Bluesky @offbrandchance.bsky.social or by email at [email protected].
Related Stories
AI News
What Do Dario Amodei And The JCHR’s AI Warnings Mean For Startups and SMEs?
19 minutes ago
AI News
Matter Venture Partners Closes $450 Million Fund II For HardTech, Robotics, And Physical AI Startups
19 minutes ago
AI News
Technologies, Math Research Get Lift from CAREER Awards
1 hour ago
AI News
Novo partners with Anthropic to speed up drug development with Claude
1 hour ago
AI News
Quebec artificial intelligence institute vandalized by multiple people
1 hour ago
AI News
Europe still needs to do more to provide real artificial intelligence reassurance
1 hour ago
AI News
'Defeated' GPT-6 Astra model spent several hours just farming potatoes after being blown up by a Creeper in Minecraft — OpenAI offering gets further than any other AI system in 141
2 hours ago
AI News
Anthropic and Microsoft Researchers Weigh Limits of Artificial Intelligence at Berkman Klein Panel
2 hours ago