TLDRocket
Sign in

Anthropic’s Text Watermarking Proves AI Companies Do Not Care at All About Writing

404 Media Jason Koebler Covered by 2 sources

Opinion — commentary, not a factual news event.

Anthropic explained how it'll watermark AI text: not hidden characters, but quietly steering word choices like "grey" vs "overcast." The backlash is about what that reveals — Anthropic treats word choice as basically meaningless.

Based on reporting by 404 Media, Jason Koebler — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

Anthropic finally explained how its promised AI text watermarking will actually work, and the mechanics are almost beside the point. What's landed the company in hot water is the reasoning behind them. In a weekend blog post, Anthropic said nothing gets added to the text and there are no hidden characters. Instead, the watermark comes from quietly changing which words Claude picks when multiple options are, in Anthropic's view, roughly equivalent.

The company's own example is the tell. Given the phrase "The weather today was cold and…," Anthropic says the next word could plausibly be "overcast" or "grey," and which one gets chosen is normally just down to a random number. Watermarking swaps that randomness for a version tied to a secret key, so Anthropic can later check whether a stretch of text matches the pattern that key would produce. Do that across enough of these so-called low-stakes choices, and you get a statistical fingerprint invisible to readers but detectable by anyone holding the key.

Anthropic insists this changes nothing about quality, creativity, or readability, citing internal testing and pointing to a Google study on a similar tool called SynthID. But the human evaluation behind that study amounts to thumbs-up and thumbs-down ratings from Gemini users on roughly 20 million responses, plus side-by-side comparisons where people apparently had no strong preference between versions. Critics — including Daring Fireball's John Gruber and journalism academic Jeff Jarvis — argue that's a weak way to judge writing, and that treating synonyms as interchangeable strips meaning out of language rather than preserving it.

The deeper issue isn't the watermark itself so much as what it exposes. Anthropic frames writing as a probabilistic exercise, something settled by an arbitrary random number generator whenever the model can't tell the difference between two words. That might be an accurate description of how a language model works. It is a much stranger claim to make about writing in general, especially from a company that has drawn scrutiny over how it sourced training material from books and other text in the first place.

Anthropic says the change is tied to new EU AI regulations, and detecting AI-generated content is a reasonable goal on its face. But the offhand way the company describes word choice as fungible, and its own comparison to code — where it says an exact output is required, unlike writing — says something about how little weight is given to the actual craft involved in choosing one word over another.

My take — AI-written commentary, not fact-checked reporting

Anthropic basically admitted, in its own blog post, that it doesn't think word choice matters — and then acted surprised when writers pushed back. Judging writing quality by Gemini thumbs-up rates on 20 million auto-generated responses is not a serious methodology, it's a rounding error dressed up as evidence. The bigger pattern here is familiar: AI companies keep treating language as interchangeable tokens because that's how their models see it, then act baffled when people who actually write for a living object to being told their choices are arbitrary.

Read more about this at: 404 Media

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.