Ant’s inclusionAI launches Ling-3.0-flash with 124B parameters and 256K context

Loose_Bank1709 · reddit · 2026-07-25

Ant Group’s inclusionAI is launching Ling-3.0-flash, a 124B-parameter sparse MoE model with about 5.1B active parameters, a 256K context window, and sub-100ms TTFT.

The model is positioned as an execution-focused node for agent workflows rather than a frontier reasoning model. It is API-only, available on OpenRouter, and offered free to use through August 3. The post also notes a hybrid reasoning mode and says the trade-offs are weaker obscure-domain knowledge and no native multimodal support.

Related event: AntLingAGI Launches Ling-3.0-flash: 124B MoE for Production Agents(11 posts)→

Original post →

More from coding & agent

coding & agent channel →