DeepSeek V4-Flash Cuts Costs 100x, Sparking AI Economics Debate

The DeepSeek-V4-Flash-0731 model has recently sparked widespread discussion regarding the true inference costs of large models. According to Artificial Analysis, this model scored close to Fable 5 on the Terminal-Bench 2.1 evaluation, with a total cost for completing the same benchmark tasks being only 1/105 of the latter. This data has driven industry attention toward shifting AI economic metrics from "cost per token" to "cost per task completed," though it has also faced skepticism regarding its practical application stability and potentially misleading billing.

Confirmed

Unconfirmed

Why It Matters

2026-08-01 ~ 2026-08-03 · 11 related posts

Primary sources

1 near-duplicate retellings: mustafamhus