Meituan Open-Sources Trillion-Param LongCat-2.0
美团技术团队 · wechat · 2026-07-09
Meituan's tech team announced the open-sourcing of LongCat-2.0. The model has a total of 1.6T parameters with an average activation of about 48B, and is available in multiple precision formats including BF16, FP8, and INT8. The team stated that the model has completed inference on a 50,000-card domestic compute cluster, focusing heavily on real-world Agentic Coding tasks, with model weights and inference code released simultaneously.
The post details LongCat-2.0's optimizations for domestic chip adaptation, KV-cache transmission, PD separated deployment, long-context processing, and sparse attention. It also covers post-training strategies for execution, reasoning, and interaction, emphasizing the goal of making a reproducible inference stack accessible to more domestic chips and existing compute resources.
Related event: Meituan Open-Sources LongCat-2.0, a Trillion-Parameter MoE Model(2 posts)→
More from Infra
- OpenRouter agents now out-consume humans as AI usage arrives in three waves — AccBalanced · 2026-09-11
- Nvidia Is Now Core to Every Major Robotaxi Stack at Commercial Scale — pdamodaran · 2026-09-11
- 12 KV Cache Reduction Techniques Every AI Engineer Should Understand, Explained — blaizedsouza · 2026-09-11
- The shadow GPU capacity market is formalizing, with Meta selling excess compute to outside buyers — DavidLinthicum · 2026-09-11
- Engram's random reads don't suit SSDs; CPU-memory over NVLink could serve all 72 GPUs — bookwormengr · 2026-09-11
- RunningHub open-sources H3Lightning, speeding up MiniMax H3 video generation 12x — 智东西 · 2026-09-11