Why AI text still reads like a bot: ICLR paper cuts slop by 90% via inference-time bans

ziv_ravid · x · 2026-09-01

The author notes a paradox: in 2026 agents run for days, ship PRs, and saturate every benchmark, yet model-written text is still identifiable within two sentences — phrases like "It's not X, it's Y" and "Let's delve" persist.

Three causes:

Their ICLR 2026 Antislop paper (with Sam Paech, Judah Goldfeder, and Allen Roush) builds a per-model slop fingerprint by comparing output against pre-2022 human writing, bans those strings at inference with a backtracking sampler, then bakes it in with FTPO — a preference method acting on single tokens in logit space — achieving about 90% less slop. Even so, AI-generated text remains easy to spot.

Original post →

More from Models

Models channel →