New paper scales local MoE by merging thousands of LoRAs at inference time

s_scardapane · x · 2026-07-29

Local MoE uses model merging to approximate test-time training

The paper proposes Test-Time Model Merging (TTMM), a way to scale Mixture-of-Experts models to far more experts with almost no test-time overhead.

Original post →

More from Research

Research channel →