DeepSeek's new model shows abnormal cache write costs, beating GPT without them
teortaxesTex · x · 2026-07-31
A developer reported that the API cost for DeepSeek's new model (suspected v4 flash 0731) in Artificial Analysis benchmarks shows an abnormally high 'cache write fee'.
If this unexplained cache creation cost is excluded, the model's performance-to-cost Pareto frontier would outperform GPT 5.6 luna.
Related event: DeepSeek-V4-Flash Benchmarks Impress but Face Pricing Controversy(4 posts)→
More from Models
- Luna Model Prices Slashed by 80%, Ushering in Cheap Intelligence Era — teortaxesTex · 2026-07-31
- Rumor: DeepSeek-V4 Performance to Rival Claude Opus 4.8 with Disruptive Pricing — teortaxesTex · 2026-07-31
- Deep20Bench tests LLM strategy via 'Twenty Questions': Opus 5 and Kimi K3 lead the pack — wauwau0977 · 2026-07-31
- DeepSeek-V4-Flash surpasses V4-Pro-Preview in latest benchmarks — Outside-Risk-8912 · 2026-07-31
- DeepSeek-V4-Flash API Public Beta: Agent Capabilities Boosted Purely via Post-Training — 新智元 · 2026-07-31
- DeepSeek-V4-Flash-0731 far surpasses DeepSeek-V4-Pro-Preview in benchmarks — SnooBunnies8392 · 2026-07-31