DeepSeek v4 Flash: How Does a 284B Model Offer Such Low Pricing?

ihatebeinganonymous · reddit · 2026-07-06

Users noticed that DeepSeek v4 Flash (284B total parameters) is priced lower on platforms like OpenRouter than the 27B Qwen model—less than a tenth of its parameter size. The core question is whether the MoE sparse activation architecture significantly reduces actual inference costs, or if providers are engaging in strategic price dumping. This highlights the structural economic advantages of MoE models in inference.

Original post →

More from Models

Models channel →