Visual guide: How AI text watermarking works by hiding in choices
ArtificialOther · x · 2026-08-17
This article provides a visual guide to the inner workings of AI text watermarking.
Core Mechanism:
- Not hidden in characters: Text lacks pixels for data storage. Instead, watermarks exist within the decision-making process of word selection.
- How it works: When generating text, a model chooses from a list of probable words (weighted dice). A secret key subtly biases the weights towards specific options, embedding a statistical pattern without altering the text's meaning.
- Industry Adoption: Google has watermarked text in the Gemini app since 2024, and Claude models have applied model-level watermarking since August 2026.
More from Safety
- Investigation: Rare Books Tracked to Amazon Facility for Scanning and Destruction for AI Training — SatelliteNetSec · 2026-08-17
- Anthropic accused of contradicting stance on AI regulation — neil_chilson · 2026-08-17
- Anthropic Reportedly Opposed Thune-Klobuchar AI Proposal — neil_chilson · 2026-08-17
- Tech giants fight back against AI-generated slop — nordicinst · 2026-08-17
- Gary Marcus criticizes OpenAI for dissolving three safety teams in two years — GaryMarcus · 2026-08-17
- AirTag Tracking Confirms Rare Books Ship to Amazon AI Training Facility — Simon Willison · 2026-08-17