A chart comparing AI detectors says most are barely better than a coin toss
burkov · x · 2026-07-23
The post argues that current AI-generated content detectors are close to useless when the task is only to decide between two classes: AI-generated or not AI-generated.
The chart compares multiple detectors such as Check for AI, Compilatio, Content at Scale, Crossplag, DetectGPT, GPT Zero, OpenAI Text Classifier, Turnitin, Writer, and Zero GPT. Most scores cluster around roughly coin-flip territory, with several tools in the 50%–70% range and a few slightly higher, but none looking reliably strong.
The takeaway is that detector accuracy remains poor enough that the author treats the problem as barely better than random guessing.
Related event: AI Content Detectors No Better Than a Coin Toss(2 posts)→
More from Research
- Nature paper images cellular activity across all organs, revealing body-wide circuits — arjunrajlab · 2026-09-11
- SignNet 1M Dataset Released for Sign Language Research — ducha_aiki · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- InFlux++ Method Released — ducha_aiki · 2026-09-11
- Skyfall GS Uses Flux to Refine Gaussian Splatting, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11
- Could 10k agents discover learning methods beyond backprop, or just tweak existing ones? — SeunghyunSEO7 · 2026-09-11