Long-context models still copy irrelevant text, and a new reward lifts accuracy by up to 4.6 points

rohanpaul_ai · x · 2026-07-23

Long-context models still waste tokens by copying irrelevant text, and a new reward fixes it

A paper on long-context reasoning finds that even strong models often fall into repetitive copying: they copy large chunks of the prompt into their internal reasoning instead of focusing on the key evidence.

The paper’s takeaway is that long-context reasoning still depends on precise grounding, not just bigger context windows.

Related event: Study Reveals Long-Context Models Waste Tokens Copying Text(2 posts)→

Original post →

More from Research

Research channel →