Why text watermarking is hard: expert explains discrete data challenge and AI Act implications

antoine_chaffin · x · 2026-08-11

AI researcher @antoinechaffin discusses the difficulty of watermarking LLM outputs: text is discrete, so you can't simply add noise like with continuous data (images, videos, audio). This is similar to the problem with text GANs. He notes that while you can't freely navigate a continuous space, there are still methods, but exposing encryption keys or detection tools carries risks. He also recalls discussions around the AI Act and recommends Meta watermarking expert @pierrefdz.

Related event: Text Watermarking Challenges Spotlighted: Discrete Data Hurdles and AI Act Boost(2 posts)→

Original post →

More from Safety

Safety channel →