DeepSeek's price hike lands Sunday: V4-Flash peak output jumps ~4.7x
NeuralNomad87 · reddit · 2026-08-17
DeepSeek's new pricing takes effect Sunday the 16th at 16:00 UTC. Peak-hour output on V4-Flash goes from $0.28 to $1.32 per million tokens (4.7x), and increases across the V4 line range from 50% to over 1,100% depending on model, input vs output, and time of day.
The author notes DeepSeek stays cheaper than most frontier APIs — this isn't a "DeepSeek is over" post — but many picked it precisely because the price made volume-heavy workloads viable: batch jobs, multi-call agent loops. If a workflow only worked at $0.28, that's when you find out.
The thread asks production users for concrete numbers: does the new pricing change your architecture? Has anyone migrated to another cheap hosted option and honestly measured quality differences? And for going local, at what monthly volume does it make sense once hardware is included?
Related event: DeepSeek Hikes V4 API Prices, Unveils Peak/Off-Peak Pricing(4 posts)→
More from Models
- EU firms may use Chinese open models via "jurisdictional wrapper" — teortaxesTex · 2026-08-17
- empero-ai's distilled Qwen3.8-9B trends on Hugging Face — empero-ai · 2026-08-17
- Dev Reports Cache Misses on gpt5.6-sol During Slow Tool Calls — lucasmeijer · 2026-08-17
- Test shows Sol and GLM-5.3 catch critical code flaws; Claude misses them — morgymcg · 2026-08-17
- DeepSeek v4 Pro Review: Matches Flash on Most Evals, Raises Scaling Questions — teortaxesTex · 2026-08-17
- Qwen 3.8 scores high on WeirdML but uses many reasoning tokens — teortaxesTex · 2026-08-17