Anthropic's Claude watermark may be a new text-marking method, clues suggest

gaganghotra_ · x · 2026-08-14

Anthropic has revealed enough about its Claude watermark to suggest it may be a novel text-marking method. Based on official announcements and transparency page updates, the watermark has six qualities: embedded in generated text, imperceptible, does not alter meaning/quality/readability, generated at model level, detectable after editing, and detectable by users and third parties. The transparency page also mentions collaboration with academia, hinting at university research origins. Several papers align, with one closely matching.

Related event: Anthropic starts watermarking Claude outputs, likely built on SynthID(3 posts)→

Original post →

More from Safety

Safety channel →