Looped LM paper: 1.6B model matches full-cache baseline with 3x smaller KV cache

rupspace · x · 2026-10-11

A new arXiv paper, Towards Looped Models Done Right, Part II: Rethinking at Fixed Points, shows that shaping fixed points in looped language models lets a 1.6B-parameter model with a 3x smaller key-value cache match the downstream accuracy of a full-cache fixed-depth baseline.

Original post →

More from Research

Research channel →