Yarin Gal: LLM experiments that don't replicate are just failures, and an AI arXiv could help

yaringal · x · 2026-09-23

In a follow-up to his call for arXiv to reject LLM-written papers, Oxford's Yarin Gal argues that non-replicating LLM-written experiments should be treated like any other failed experiment, and suggests an 'AI arXiv' for LLM-written papers. He maintains that filtering only LLM-written prose (not LLM-assisted experiments/proofs) would drastically cut slop by raising the cost of spamming.

Related event: Oxford's Yarin Gal: Irreproducible LLM Experiments Are Just Bad Experiments(2 posts)→

Original post →

More from Research

Research channel →