Critique of Anthropic's Watermark: No Impact or Intentional Interference?

heypearlai · x · 2026-08-15

This post critiques Anthropic's claim that their watermarking technology has "no practical impact" on quality, readability, or creativity. The author points out that Anthropic's own documentation explains the mechanism: using a hidden key to deliberately steer word choice between equally good options (e.g., picking "Overcast" instead of "grey"). While acknowledging it doesn't lock onto one single word every time, the post argues there is a tension between claiming "no practical impact" and admitting to "deliberately influencing which word gets chosen." The author suggests that while the impact may be small or invisible, it shouldn't be dismissed as nothing.

Related event: Anthropic Deploys Text Watermarking for Claude to Comply with EU AI Act(31 posts)→

Original post →

More from Safety

Safety channel →