Study Reveals Long-Context Models Waste Tokens Copying Text

A new study reveals that long-context reasoning models waste computational tokens by copying irrelevant input text into their reasoning chains. To address this inefficiency, researchers introduced a new reward mechanism that improves model performance by up to 4.6 points.

2026-07-23 ~ 2026-07-23 · 2 related posts