NeurIPS reviewer says AI catches 10x more issues, but can still praise LLM slop
mariyaivasileva · x · 2026-07-27
A NeurIPS associate chair says well-executed AI reviews can beat human reviews by a wide margin on concrete issues:
- far more mathematical mistakes
- missing citations
- unsupported novelty claims
- experimental flaws
- typos, grammar, and inconsistencies
But a separate example shows the downside: an obviously LLM-generated paper received a long, rigorous AI review that still praised its “excellent writing,” while shorter human reviews correctly called out the bad prose.
Related event: NeurIPS Reviewers: AI Catches 10x More Issues(2 posts)→
More from Research
- PIRL adds closed-loop verification to RL post-training and improves accuracy on reasoning and code tasks — This_Ad9834 · 2026-07-28
- Vstone unveils a 97 kg mobile dual-arm humanoid robot for factory and logistics work — CyberRobooo · 2026-07-28
- A 200-patient synthetic table stayed unique after removing all identifiers — MaziyarPanahi · 2026-07-28
- GitHub repo reproduces Patch Policy for embodied control with DINOv2 features — k7agar · 2026-07-28
- Simons Institute panel asks how researchers should adapt to automation — ceciletamura · 2026-07-28
- A research agent works better when it can stop instead of forcing an answer — Harshit-24 · 2026-07-28