Fact-Checking the NeurIPS-Pangram Story: v3.3.2 Flagged 14.9% of Human-Written Reviews as AI

tughanbulut · reddit · 2026-09-26

The author revisits the NeurIPS Position Track / Pangram controversy after decisions came out, correcting several claims from an earlier post against primary sources.

What holds: Pangram's v4 report shows v3.3.2, the exact version NeurIPS used, flagged 14.9% (easy) and 4.5% (hard) of human-written, AI-polished peer reviews as fully "AI" — worse than both v3.0 and v4 — while the track policy explicitly allowed polishing. An ICML 2026 paper says per-window false-positive rates shouldn't be extrapolated to whole papers in either direction. The clearance status of the 123 conditional papers remains undisclosed.

Original post →

More from Safety

Safety channel →