Ling-3.0-flash offers 124B total parameters with just 5.1B active per token
truecakesnake · reddit · 2026-07-25
Ling-3.0-flash is being offered free in Nous Research’s portal for one week, with the post highlighting its unusual MoE shape: 124B total parameters but only 5.1B active per token.
The thread argues that this low-active configuration is attractive for agent workloads like coding, search, research, and tool use, especially on bandwidth-limited hardware such as Macs, Strix Halo systems, DGX Spark boxes, and large-RAM CPU machines.
Related event: Ling-3.0-flash Released: Targeting Agents with Long Free Access(10 posts)→
More from Infra
- AI demand is pushing DRAM prices up 171.8% and relief may not come until 2028 — 量子位 · 2026-07-25
- Reproducible backends could make GPU and CPU profiling portable, but the business model is unclear — omojumiller · 2026-07-25
- A practical document-parser playbook says layout and scan quality matter more than rankings — emmettvance · 2026-07-25
- NURL nears v1.0 with a single-binary language built for LLM workflows — AdhesivenessHappy873 · 2026-07-25
- Anthropic says it has supply deals with Samsung Electronics and SK hynix — dejavucoder · 2026-07-25
- OpenAI users report `primaryapi_server_error` after login failures — HaselnuesseTo · 2026-07-25