DeepSeek cuts flash pricing up to 60% as v4.1 flash quietly goes live
量子位 · wechat · 2026-09-09
Per QbitAI, DeepSeek announced new price cuts for its flash series: per-million-token off-peak cache pricing drops from ¥0.05 to ¥0.02 (-60%), off-peak input from ¥1.5 to ¥1 (-33%), and off-peak output from ¥4.5 to ¥4 (-11%). Peak-hour rates are 2x off-peak.
A new DeepSeek v4.1 flash model has also quietly launched — callable via the model name deepseek-v4.1-flash-expires-on-0910 in the API.
More from Models
- ValsAI launches RSI Index, first third-party benchmark measuring how close AI is to self-improvement — JenniferHli · 2026-09-11
- Devin's New Model Verdict: Not a Benchmaxxer, a 'Killer Execution Model' at $20/Month — brandon_galang · 2026-09-11
- Business Insider Asked ChatGPT, Gemini, Claude and Grok How AI Could End Humanity — coinfanking · 2026-09-11
- Claims resurface that Moonshot's Kimi distilled from Claude raw CoTs — xuanalogue · 2026-09-11
- User switches back to GPT-5.6 Sol: barely uses quota and feels faster — CtrlAltDwayne · 2026-09-11
- Dev opinion: model differences shrink in a good harness; Grok 4.6 is good enough — gnukeith · 2026-09-11