Ant's Ling 3.0 Tiny Activates Only 1.3B of 7.9B Params for Agents
truecakesnake · reddit · 2026-08-07
InclusionAI, a lab under Ant Group, has released Ling 3.0 Tiny. The model has a total of 7.9B parameters but only activates about 1.3B per token, specifically targeting multi-turn agent workflows.
Key Specs & Availability:
- Supports 256K context length with up to 32K output tokens.
- Features native function calling, prompt caching, and a toggle between thinking and instant modes to reduce latency.
- Closed Source: Currently offered only as a hosted API via Vercel, OpenRouter, etc. No weights are available for self-hosting.
Industry Implications: As activated parameter counts drop, the cost of the middle-layer models in agent architectures approaches zero, turning model selection from an architectural decision into a cheap commodity purchase.
Related event: AntLingAGI Releases Ling-3.0-tiny Native Hybrid Reasoning Model(2 posts)→
More from coding & agent
- Open-Sourcing Mandate: A Financial Stack Giving AI Agents Economic Autonomy — RichardsonDx · 2026-08-07
- Pydantic Creator Prefers Claude Code, Leaves OpenAI Free Tokens Unused Due to Poor UX — samuelcolvin · 2026-08-07
- Mediabunny optimizes large file streaming with on-demand downloads and memory control — Vjeux · 2026-08-07
- The Future is AI-to-AI Loops: Agents Will Negotiate APIs and Report Bugs — DanielLockyer · 2026-08-07
- Developers Frustrated by Claude Code Harness, Leaving Free Tokens Unused — intellectronica · 2026-08-07
- AI Engineering is Shifting to Systems Design: Context and Routing Take Center Stage — Deep_Ladder_4679 · 2026-08-07