DeepSeek-V4-Flash Hits Cost-Performance Frontier on Agent Arena

infwinston · x · 2026-08-05

DeepSeek-V4-Flash has achieved a new breakthrough in cost-performance for agentic tasks.

According to real-world data from Agent Arena, the model (High mode) delivers a median cost of $0.024 per task. It outperforms GPT-5.6 Luna (xHigh at $0.026) and delivers a positive net improvement at the lowest price point on the chart, sitting just to the right of DeepSeek-V4-Pro (Thinking at $0.020).

The cost is calculated based on real Agent Mode usage, factoring in token consumption, model pricing, and cache hits/misses.

Related event: DeepSeek-V4-Flash Tops Cost-Efficiency with Ultra-Low Running Costs(4 posts)→

Original post →

More from Models

Models channel →