DeepSeek v4 Flash: How Does a 284B Model Offer Such Low Pricing?
ihatebeinganonymous · reddit · 2026-07-06
Users noticed that DeepSeek v4 Flash (284B total parameters) is priced lower on platforms like OpenRouter than the 27B Qwen model—less than a tenth of its parameter size. The core question is whether the MoE sparse activation architecture significantly reduces actual inference costs, or if providers are engaging in strategic price dumping. This highlights the structural economic advantages of MoE models in inference.
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11