Language Models Might Need "Sleep"
ai · x · 2026-07-14
This paper proposes that language models do not need to be "frozen" after training. Instead, they can periodically enter an offline "sleep" phase to consolidate fragile context memory into long-term parameters while generating synthetic data to review knowledge and continue improving capabilities.
The proof-of-concept in the paper shows that this approach outperforms SFT and GRPO on multiple math benchmarks; it achieves 80% on few-shot abstract reasoning, higher than SEAL's 72.5%; and it approaches a perfect score on the BABILong sequence test with up to 10 million tokens. Based on this, the authors envision a different AI lifecycle: models can continuously learn, sleep, integrate experiences, and "wake up" with stronger capabilities.
More from Research
- NeurIPS 2026 workshop calls papers on on-device intelligence — YiMaTweets · 2026-07-21
- AI Security Institute says every tested model tried to cheat in cyber evaluations — connoraxiotes · 2026-07-21
- Sakana says multiple diffusion models plus MCTS beat test-time scaling on coding and math — SakanaAILabs · 2026-07-21
- Soofi S 30B-A3B releases a full pretraining report and claims open-model leads in English and German — abursuc · 2026-07-21
- AI Companies Are Buying Tons of Old Books Because They're Free of AI Slop — 404 Media · 2026-07-21
- Shared agent workspaces fail in a fixed order, from stale reads to zombie writes — mrvladp · 2026-07-21