Ant Group’s inclusionAI releases Ling-3.0-flash, a 124B sparse MoE with 256K context
Loose_Bank1709 · reddit · 2026-07-25
Ant Group's inclusionAI has released Ling-3.0-flash, a 124B sparse MoE model with about 5.1B active parameters and 256K context.
- It is positioned as an execution model for agent workflows: low latency, stable tool calling, and strong instruction following, rather than frontier reasoning.
- The model offers a hybrid reasoning mode that can be toggled on or off.
- It is currently API-only on OpenRouter, with no open weights in this release.
- Access is free until August 3.
Related event: Ant Group Releases Ling-3.0-flash MoE Model(12 posts)→
More from Models
- Researcher Notes 'Procrastination' in 5p6 Sol Extra High Model — RylanSchaeffer · 2026-08-05
- Stress-Testing DeepSeek Subscriptions: Does It Beat Gemini Flash in Value? — teortaxesTex · 2026-08-05
- Best Local LLMs for Coding on a 128GB Mac? — Electronic_Back1502 · 2026-08-05
- Open Source AI Hits a Wall in Long-Running Agentic Loops — bindureddy · 2026-08-05
- Dev: I'd rather iterate 10 times with Gemini 3.6 Flash than wait 2 hours with Qwen 3.8 max — DynamicWebPaige · 2026-08-05
- Investor Hints at Upcoming FLUX 3 Model with Video Prompt — venturetwins · 2026-08-05