How AI text watermarking embeds hidden signatures in word choices
AI text watermarking works by subtly influencing which words a language model selects during text generation, without altering the model's underlying weights or making the output appear unusual. A secret key combined with preceding tokens assigns hidden scores to candidate words, and a method called tournament sampling uses these scores to determine which word is chosen. The resulting text looks entirely normal but carries a detectable pattern that persists even when copied across platforms. Google DeepMind introduced the underlying technique, SynthID-Text, in a 2024 Nature publication, and Anthropic adopted a version of it for its Claude model in August 2026. Anthropic stated the move is intended to satisfy transparency requirements under the EU AI Act.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in