DeepSeek V4 Flash Ties Gemini 3.6 Flash in Intelligence at 1/30th the Output Cost
alejandroll10 · x · 2026-07-31
DeepSeek officially launched the V4-Flash API in public beta, featuring massively upgraded Agent capabilities, native Responses API support, and Codex adaptation.
Third-party benchmarks reveal that V4-Flash ties Gemini 3.6 Flash on the intelligence index but is 30x cheaper for output ($0.28 vs $7.50 per million tokens). With open weights and extreme cost-efficiency, it challenges the premium pricing of closed-source models.
Related event: DeepSeek-V4-Flash Enters Public Beta with Enhanced Agent Capabilities(32 posts)→
More from Models
- Kimi K3 Outperforms Claude Opus and Shows Superior Context Efficiency — casper_hansen_ · 2026-07-31
- More GB300s Online on LightningAI; Pangram 4 Hits 99% AI Text Detection Accuracy — LightningAI · 2026-07-31
- ChatGPT still cites old URL 2 weeks after redirect, search updated in 24h — lilyraynyc · 2026-07-31
- Specific Prompt Bypasses Guardrails to Unlock Claude Opus Base Model — paul_cal · 2026-07-31
- OpenAI Price Cuts and DeepSeek Update Expose Anthropic's Model Pricing Dilemma — kimmonismus · 2026-07-31
- DeepSeek-V4-Flash Runs at 400 tps for Just $10/hour on Inference Endpoints — ben_burtenshaw · 2026-07-31