TLDRocket
Sign in

How Claude's text watermarking works

Anthropic Covered by 2 sources

Claude is getting text watermarks so its output can be spotted later. Anthropic says readers won’t notice, but EU rules now demand the mark.

Based on reporting by Anthropic — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

Anthropic says future Claude models will start generating text with a watermark, a hidden pattern that lets someone estimate whether Claude helped write it. The company says the change is about EU compliance, not product tinkering. And it insists the reading experience won’t change at all.

The trick is in the model’s random choices. Large language models pick words one at a time, and many of those choices are basically interchangeable: overcast or grey, for example. Anthropic’s watermark nudges those low-stakes selections using a key, so the sequence of words ends up consistent in a way that can be checked later. To the reader, though, the text should look identical.

That matters because Anthropic says the watermark adds nothing visible to the text, no hidden characters, no extra tokens, and no extra cost. It also says the watermark carries no identifying information and can’t be traced back to a person, organization, or chat. The company says its internal testing found no hit to quality, creativity, or readability, and it points to Google DeepMind’s SynthID-Text work, which reported no statistically significant difference in ratings.

The watermark is not a magic lie detector. It can say whether text was likely partly written by Claude, but it cannot prove human authorship, identify a different AI, or work well on short samples. It also gets thinner in places where the wording has to be exact, like factual passages, proofreading, and much of code. In those cases there are fewer choices to hide a pattern in.

Anthropic says the rollout is tied to the EU AI Act, which requires providers serving the market to mark AI-generated content as of August 2. The company says it is applying watermarking globally at launch because it doesn’t yet have a durable way to limit it by region. It also says a watermark detection API is coming soon, while files such as .png, .jpg, and .svg will use C2PA content credentials instead of a text watermark.

My take — AI-written commentary, not fact-checked reporting

This is the right kind of AI transparency: boring, practical, and not pretending a watermark is a truth machine. The real tell here is that the industry only seems capable of sharing standards when Europe makes it mandatory, which is a very on-brand way for Big AI to discover conscience.

Read more about this at: Anthropic

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.