Ant’s inclusionAI launches Ling-3.0-flash with 124B parameters and 256K context
Loose_Bank1709 · reddit · 2026-07-25
Ant Group’s inclusionAI is launching Ling-3.0-flash, a 124B-parameter sparse MoE model with about 5.1B active parameters, a 256K context window, and sub-100ms TTFT.
The model is positioned as an execution-focused node for agent workflows rather than a frontier reasoning model. It is API-only, available on OpenRouter, and offered free to use through August 3. The post also notes a hybrid reasoning mode and says the trade-offs are weaker obscure-domain knowledge and no native multimodal support.
Related event: Ant Group Releases Ling-3.0-flash MoE Model(12 posts)→
More from coding & agent
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- Codex tip: use Sol with Astra and Luna sub-agents to save usage — pvncher · 2026-09-11
- agents-best-practices: a provider-neutral Agent Skill for designing and auditing agentic harnesses — tom_doerr · 2026-09-11
- Cognition's SWE-2 uses a KKT duality argument in RL to shift the effort Pareto curve — YouJiacheng · 2026-09-11
- First-ever Three.js Conference lands in Paris, with a panel on AI-shortened design workflows — OdinLovis · 2026-09-11
- Agile co-author Ron Jeffries publishes 'Resist AI', urging developers to push back — mborch · 2026-09-11