Periscope Extends Frozen LMs to Million-Token Contexts Without Training

KAUST's Periscope is a training-free inference method that lets frozen language models answer questions over documents far beyond their context window, enabling a 27B model on a single 80GB GPU to process about 4.5 million tokens.

2026-10-06 ~ 2026-10-06 · 2 related posts