Ant’s inclusionAI launches Ling-3.0-flash with 124B parameters and 256K context
Loose_Bank1709 · reddit · 2026-07-25
Ant Group’s inclusionAI is launching Ling-3.0-flash, a 124B-parameter sparse MoE model with about 5.1B active parameters, a 256K context window, and sub-100ms TTFT.
The model is positioned as an execution-focused node for agent workflows rather than a frontier reasoning model. It is API-only, available on OpenRouter, and offered free to use through August 3. The post also notes a hybrid reasoning mode and says the trade-offs are weaker obscure-domain knowledge and no native multimodal support.
Related event: AntLingAGI Launches Ling-3.0-flash: 124B MoE for Production Agents(11 posts)→
More from coding & agent
- How to handle captchas when running Claude Code and Codex automation — dry_Relationship2007 · 2026-07-25
- How are people seeding agent sandboxes with realistic data? — neal_lathia · 2026-07-25
- How teams stop autonomous agents from burning API credits in production — Odd_Weather_3289 · 2026-07-25
- Developer submits Linux + Codex support PR after using the tool daily on Linux — cneuralnetwork · 2026-07-25
- Free MCP connector brings Amazon CPG pricing intelligence into AI workflows — modelcontextprotocol · 2026-07-25
- Todoist MCP server lets AI assistants manage tasks, projects, and labels by chat — modelcontextprotocol · 2026-07-25