How AI text watermarking works

How AI text watermarking works

A visual, interactive explainer of how statistical watermarks hide inside AI-generated text, and what erases them. It covers the core mechanism (secret-keyed nudging of word choices), how detection works via counting, the effect of editing, and practical implications. Includes demos and references to real systems like Google's SynthID and Anthropic's Claude.

They're invisible, they survive copying, and they work because they don't live in the characters at all. They live in the choices between them.
  1. wpasc

    What's wild (imo) is pretty much everyone I talk to/read from (anecdata) HATES the way claude writes. I see it in the comments on Hacker News, hear about it in discussions with my colleagues, and talk about it with my non tech family. it's over the top bad. now it seems like these quirks will now be enforced in some weird way to meet the watermarking rules?

  2. gizmo686

    I could see this being useful in a world with a few AI providers. However, in a world of commodity AI models, can simply use a model from an AI provider that does not watermark. Or download any open source model and run it themselves [0].

    The only practical use I can see for this in the world we actually live in is to prevent model collapse. Most people using AI don't care if people training future AI ignore them, so would have no incentive to switch to providers that do not watermark. Of course, this disencetivises all if the pro-social applications of this technology, and risks giving the big providers a monopoly on "known human" data, which has serious antitrust implications.

    [0] Note that the watermark is not inherent to the model itself, but rather how the model is run. So this teqnique cannot be used by people providing open-weight models. It would need to be used by those actually running the models.

  3. guessmyname

    I almost never copy & paste AI-generated text, I almost always transcribe it by hand, which in turn forces me to read what the LLM generated and gives me the opportunity to replace words as I go. This obviously doesn’t scale, especially if your impact is measured by the number of software features you implement, but for more experienced engineers (Staff, Principal, and above) who are usually evaluated on the success of company-wide initiatives, I think this is the best course of action.

  4. rrgok

    But text watermarking is not possible in coding, right?

  5. throwatdem12311

    New job idea: have a human reword/summarize and manually input/transcribe AI output to remove the watermarking. They use “tools” like dictionaries and thesaurus’ in order to sufficiently change the text so that it doesn’t fit within AI distribution anymore. Humans that can write significantly “organic” text will be able to make lucrative careers out of it.

More from this day

2026-08-14