AI Detectors Crumble Against Fine-Tuned Models: Pangram Falls to 3% Accuracy, GPTZero to 0%
Dr_WhoDo · reddit · 2026-09-16
A Reddit post exposes a major blind spot in AI detection tools.
- Pangram and GPTZero claim 97% and 91% accuracy against style-prompted LLMs with near-zero false positives
- But against models fine-tuned to mimic specific authors or bodies of work, Pangram's accuracy collapses to 3% and GPTZero's to 0%
- Commercially available models can already be fine-tuned to any author's style, and detectors have no answer
The author also cites the Stony Brook / Columbia Law / MIT study Readers Prefer Outputs of AI Trained on Copyrighted Books over Expert Human Writers — readers actually prefer the fine-tuned AI prose. Conclusion: detectors' headline numbers only hold against average LLM output, and vendors are masking a fundamental flaw in their solutions.
More from Models
- Qwen 27B on a single RTX 5090 generates 8-scene animation from one prompt — Acceptable-Object390 · 2026-09-16
- Qwen 27B on a single RTX 5090 generates 8-scene animation from one prompt — Acceptable-Object390 · 2026-09-16
- New paper lifts LLM memory retention 28x to 34.9% across 100 sequential tasks — DanielKhashabi · 2026-09-16
- Claude usage bug unexpectedly grants extra quota as user plans ASC CLI release — rudrank · 2026-09-16
- xAI Launches Grok 4.6 With Unified Multi-Modal API, Sub-200ms Latency — Jasonio · 2026-09-16
- Mozilla Report: China-US AI Model Capability Gap Narrows to 4.4 Months — External_Mood4719 · 2026-09-16