Research finds AI watermarking like SynthID-Text shifts LLM behavior and can weaken safety guardrails

Ars Technica AI · rss · 2026-09-18

To comply with new EU law, AI platforms are deploying watermarking; Anthropic says future Claude models will use Google's open-source SynthID-Text, which uses a secret key to subtly alter token selection for provenance verification.

Original post →

More from Safety

Safety channel →