100M output tokens for $60: DeepSeek off-peak pricing undercuts Opus 5 by 40x
airesearch12 · x · 2026-09-10
A price comparison of first-party API rates: 100M billed output tokens cost $2,500 on Opus 5, $2,000 on Sol, but just $60 on DeepSeek off-peak ($120 peak) — a 40x gap. DeepSeek's docs show deepseek-flash (V4.1-Flash) with 1M context, 384K max output, peak/off-peak pricing (off-peak at half rate) and 2500 concurrency; deepseek-v4-pro bills up to $1.32/$3.96 per 1M input/output at peak. DeepSeek says V4.1 Flash now comprehensively beats V4 Pro and will retire V4 Pro: from Sept 14, v4-pro requests route to V4.1 Flash at Flash pricing.
More from Infra
- mlx-omarchy 0.4.0: run MLX on the Apple GPU under Linux, no Metal needed — alexcovo_eth · 2026-09-10
- Gemini V4.1 pretraining estimated at ~5e24 FLOPs, two weeks on 4K B300s — teortaxesTex · 2026-09-10
- 4x 3060 12GB tensor-parallel rig: is adding a 4060 Ti 16GB worth it? — Qwen30bEnjoyer · 2026-09-10
- Autonomous adds Omarchy OS to its $26,100 dual-RTX-5090 AI workstation — dee_hw · 2026-09-10
- Dev take: token demand will grow far faster than demand for top-line intelligence — willcb · 2026-09-10
- Arm lands Lenovo and ByteDance's Volcengine as first China customers for its AI server chips — pstAsiatech · 2026-09-10