How LLM Text Watermarking Works: GPTZero CTO Breaks Down the KGW Method

mark_k · x · 2026-08-11

The CTO of GPTZero explains how frontier labs like Anthropic, Google, and OpenAI are building text watermarking. Most efficient methods rely on the KGW algorithm:

While computationally cheap, this approach has known vulnerabilities and can potentially be defeated.

Related event: Anthropic Adds Invisible Watermarks to Claude Outputs, EU AI Act Compliance Sparks Debate(114 posts)→

Original post →

More from Safety

Safety channel →