Ant's Ling-3.0-flash Activates Only 5.1B Params: Architecture and Cost Analysis

jkris050 · reddit · 2026-08-05

Ant Group's lab inclusionAI open-sourced the Ling-3.0-flash model under the MIT license on Aug 4. It has 124B total parameters but only activates 5.1B (about 1/64) per token, utilizing 512 routed experts and 1 shared expert.

Architectural Highlights:

Performance & Cost Debate:

Related event: Ant's Ling-3.0-flash Open-Sourced: 124B MoE with Only 5.1B Active Parameters(5 posts)→

Original post →

More from Infra

Infra channel →