Meituan Open-Sources Trillion-Param LongCat-2.0
美团技术团队 · wechat · 2026-07-09
Meituan's tech team announced the open-sourcing of LongCat-2.0. The model has a total of 1.6T parameters with an average activation of about 48B, and is available in multiple precision formats including BF16, FP8, and INT8. The team stated that the model has completed inference on a 50,000-card domestic compute cluster, focusing heavily on real-world Agentic Coding tasks, with model weights and inference code released simultaneously. The post details LongCat-2.0's optimizations for domestic chip adaptation, KV-cache transmission, PD separated deployment, long-context processing, and sparse attention. It also covers post-training strategies for execution, reasoning, and interaction, emphasizing the goal of making a reproducible inference stack accessible to more domestic chips and existing compute resources.
Related event: Meituan Open-Sources LongCat-2.0, a Trillion-Parameter MoE Model(2 posts)→
More from Infra
- AI performance is increasingly limited by materials science, not just compute — nordicinst · 2026-07-21
- A GLM-5.2 inference debate asks how 750B parameters can exceed 1 token per second — francoisfleuret · 2026-07-21
- Why vector databases slow AI agents down after constant writes — PrajwalTomar_ · 2026-07-21
- Larry Fink says China is ahead in the AI energy race, citing 100 GW nuclear buildout — rohanpaul_ai · 2026-07-21
- Local AI may pay back in 6–7 years and cut long-term costs by 30–40% — DavidLinthicum · 2026-07-21
- TSMC reportedly plans up to 10% chipmaking price hikes in 2027 — kimmonismus · 2026-07-21