Ant Group Releases Ling-3.0-flash: 124B Parameters with Just 5.1B Active

SonglinYang4 · x · 2026-07-25

Ant Group's Ling team has released Ling-3.0-flash, a hybrid-reasoning MoE model built for production-scale agents. The model has 124B total parameters but activates only 5.1B per token. With 1/8 of the total and 1/12 of the active parameters, it reportedly matches or beats their 1T flagship model on most benchmarks. The vLLM team also praised its announce-first, open-source-next approach, noting it provides a stable window for open-source inference projects to prepare for day-0 support.

Related event: Ant Group Releases Ling-3.0-flash MoE Model(12 posts)→

Original post →

More from Models

Models channel →