Spotlight architecture adds growing memory to LLMs with constant-cost token generation

teortaxesTex · x · 2026-10-03

Christos Tzamos's team unveiled Spotlight, a new architecture claiming LLMs can gain knowledge and capabilities without retraining: models get growing memory while token generation stays constant-work regardless of stored information, marrying attention's capacity with the linear cost of fixed-state recurrent models.

Commenter teortaxesTex notes the constant factor looks large — each access touches 9 cells, each storing a full dk×dv matrix per head — though that may be inherent to the continual-learning design.

Related event: Percepta Unveils Spotlight: Decoupling Intelligence from Ever-Growing Memory(6 posts)→

Original post →

More from Models

Models channel →