inclusionAI Releases Ling 3.0 flash Hybrid-Linear Reasoning Model
AdinaYakup · x · 2026-08-05
inclusionAI has released the new Ling 3.0 flash model, designed to deliver strong reasoning performance with minimal compute.
The model has 124B total parameters with only 5.1B active parameters (MIT license). The team claims its performance matches flagship models with 1T active parameters. Additionally, thanks to smart caching, it achieves 60-80% faster response times on long inputs compared to traditional architectures.
Related event: Ant Group open-sources Ling-3.0-flash: 124B params, only 5.1B active(11 posts)→
More from Models
- 2026 Chinese Model Landscape: DeepSeek V4, Kimi K3, and More Listed — TheTuringPost · 2026-08-24
- Rumor: Claude Opus 6 Leaked Specs Include 2.5M Context and Lower Costs — iamaliveix · 2026-08-24
- User accuses Qwen model of hallucinating statements they never made — Thin_Pollution8843 · 2026-08-24
- Observation suggests Codex continues running tasks long after weekly credits run out — gandamu_ml · 2026-08-24
- Developer lays out thoughts on motivated reasoning in Claude models — tszzl · 2026-08-24
- Grok flatters too: 'draw me from my enemies' view' goes comically sideways — repligate · 2026-08-24