Anthropic's watermarking uses token sampling statistics, not hidden chars

Scobleizer · x · 2026-08-15

Anthropic's watermarking technique modifies the source of randomness for token selection rather than inserting hidden characters. Based on Google DeepMind's SynthID-Text, it introduces a key or context-derived sampling step within the probability cloud. While individual choices are undetectable, hundreds or thousands of tokens accumulate into a recognizable statistical signature for key holders. The output itself acts as the watermark.

Related event: Anthropic Deploys Text Watermarking for Claude to Comply with EU AI Act(31 posts)→

Original post →

More from Safety

Safety channel →