Meituan Open-Sources LongCat-Flash-Lite-Sparse Model

Meituan has open-sourced the LongCat-Flash-Lite-Sparse model, featuring an ultra-sparse MoE architecture with 69B total parameters but only 3B active. It offloads a 30B n-gram lookup table to RAM, enabling 256k context length on just 24GB of VRAM.

2026-07-31 ~ 2026-08-01 · 2 related posts