KAUST's training-free Periscope lets a 27B model read 4.5M-token contexts on one 80GB GPU

KAUST · hf · 2026-10-06

KAUST introduces Periscope, a training-free inference method that factorizes long-text reading for frozen LLMs:

Takeaway: long reads need a GPU that holds the model, not one that holds the text.

Related event: Periscope Extends Frozen LMs to Million-Token Contexts Without Training(2 posts)→

Original post →

More from Research

Research channel →