Papers with Code spotlights Looped Transformer: recurrent depth decoupled from parameters
NielsRogge · x · 2026-09-02
Niels Rogge highlighted the Looped Transformer topic page on Papers with Code. The method reuses Transformer blocks across recurrent depth, decoupling effective computation depth from parameter count, enabling iterative refinement, adaptive computation, and test-time depth scaling. Collected papers span HRM-Text, Nanbeige4.2-3B (agentic capabilities in a compact model), LoopViT, SMELT scaling laws for MoE looped transformers, Gated Recurrent Transformers, DeepLoop, learned halting gates, looped models improving compositional tool calling, and more.
More from Research
- Prompt optimization may not need search trees: NPO paper shows teacher quality beats elaborate search — rohanpaul_ai · 2026-09-02
- Schmidhuber revisits award-winning NLSOM paper on economies of minds — SchmidhuberAI · 2026-09-02
- MENO hybrid physics-AI simulation cuts plasma computation from 5 days to 4 hours — bravo_abad · 2026-09-02
- Jasper AI open-sources a full cookbook to train a text-to-image model from scratch — dh7net · 2026-09-02
- 2021's CABiNet beats YOLO26-sem on UAVid: +2.7 mIoU at 3x lower latency — Naive-Explanation940 · 2026-09-02
- Research agenda proposed for 'generative cryptography': AI writing crypto protocols — DavideCrapis · 2026-09-02