AI Detectors Fail When Imitating Writing Styles
FlorianGallwitz · x · 2026-07-18
An Epoch AI study tested several AI detectors across different text types:
- For human writing, detectors rarely misidentified it as AI text.
- For directly generated LLM text, the false negative rate was extremely low, around 0%0.7%.
- However, when an LLM was prompted to mimic a real author's style, detectors failed noticeably: the false negative rate across different tools jumped to roughly 10%18%.
- In scientific writing scenarios, the false negative rate was even higher, reaching about 26%.
The chart also details sample sizes and error margins. The conclusion is that while AI detectors are good at catching "blatant AI text," their reliability drops significantly when faced with content designed to mimic specific authors.
More from Research
- SUFLECA shows NOC-based correspondence can improve CAD-to-image alignment — ducha_aiki · 2026-07-21
- OpenAI-style autonomous researchers could become real scientific collaborators — Promptmethus · 2026-07-21
- Soft Clamp cuts tool-call overuse in multi-teacher distillation, from 13.7% to 9.0% — antgroup · 2026-07-21
- ShotPlan adds learnable planning tokens for cinematic multi-shot video generation — Tele-AI · 2026-07-21
- A silicon photonic reservoir chip compensates fiber distortion in real time at 28 Gbps — bravo_abad · 2026-07-21
- A developer maps out six design rules for CLIs that humans and AI agents can both use — yujiezha · 2026-07-21