Three hours of agentic research shows frontier LLMs can evade AI detectors

maier_ak · x · 2026-08-26

A new arXiv paper, 'Beating the Style Detector', used a modern agentic-research harness to redo every experiment of an ACL 2026 study on personal-style post-editing — with the human acting only as reviewer-in-the-loop — reproducing all 7 preregistered hypotheses and the headline correlation to three decimals (r=+0.244, p<10⁻⁸, n=648).

Key findings:

Original post →

More from Research

Research channel →