Style Imitation Breaks AI Text Detectors
The Decoder · rss · 2026-07-19
Epoch AI tested three mainstream AI text detectors (Pangram, GPTZero, Originality.ai) on texts designed to "imitate an author's writing style."
Results show that this type of style imitation significantly weakens detection efficacy: up to 18% of AI-generated paragraphs went undetected; in scientific writing scenarios, the false negative rate reached as high as 48%. The article points out that this is precisely the domain where these detectors are most likely to be deployed in real-world applications.
More from Safety
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- California creates standards for independent AI auditors to verify lab safety testing — VraserX · 2026-09-11
- a16z podcast: why 2-3 person startups are absent from policy debates — a16z Podcast · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- Class action accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers — The Decoder · 2026-09-11
- MD shows buying lab media requires background checks, calling AI bioweapon doom scenarios implausible — Ghost_Pilot_MD · 2026-09-11