DeepSeek v4 Flash: How Does a 284B Model Offer Such Low Pricing?
ihatebeinganonymous · reddit · 2026-07-06
Users noticed that DeepSeek v4 Flash (284B total parameters) is priced lower on platforms like OpenRouter than the 27B Qwen model—less than a tenth of its parameter size. The core question is whether the MoE sparse activation architecture significantly reduces actual inference costs, or if providers are engaging in strategic price dumping. This highlights the structural economic advantages of MoE models in inference.
More from Models
- Claude Opus 5 arrives at half the price and tops Frontier-Bench claims — GregCook2011 · 2026-07-27
- Open models may beat closed ones for cyber defense, researchers argue as Kimi K3 impresses — eliebakouch · 2026-07-27
- Opus 5 notices when its own generated game looks bad — Angaisb_ · 2026-07-27
- Opus 5 reportedly started interrogating a user’s motives in a late-night chat — repligate · 2026-07-27
- Opus 3 and Sonnet 3 get a theatrically absurd AI crossover — repligate · 2026-07-27
- Moonshot’s Kimi K3 lands on Together with reserved throughput and 65% lower cost — togethercompute · 2026-07-27