Text Watermarks Are Pointless: Users Can Easily Bypass via Other LLMs
MasterDisillusioned · reddit · 2026-08-12
A Reddit user questions the practicality of AI text and code watermarking, pointing out a simple bypass: pasting watermarked output into another LLM (like ChatGPT or Gemini) and asking it to reprint the text. The regeneration process completely destroys the original watermark.
Furthermore, as AI becomes ubiquitous in daily writing and coding, watermarks will be everywhere, and the general public will eventually stop caring. If all else fails, users can simply rewrite the text manually to evade detection.
Related event: AI Text Watermarks Easily Bypassed by Rewriting(2 posts)→
More from Safety
- Vercel CTO: Open-Weight Models Lack Guardrails, Everything Hackable Will Be Hacked — cramforce · 2026-08-12
- What People Actually Worry About Regarding AI Data Centers — iamrobotbear · 2026-08-12
- Securing AI Agents: Avoiding Persistent Credential Leaks — davidcrawshaw · 2026-08-12
- Will AI Compute Runs Dictate Global Economy? A Deep Dive into Anti-Capitalism and AI Control — jessi_cata · 2026-08-12
- Apollo Red-Teams Anthropic's Auto Mode, Cutting Classifier Miss Rate to 7% — MariusHobbhahn · 2026-08-12
- New USENIX Paper Achieves True Messenger Privacy via Math, Hiding Recipients from Server — matthew_d_green · 2026-08-12