Baidu serves DeepSeek V4 Flash at crazy fast speeds and low prices

NielsRogge · x · 2026-08-19

Users have observed that Baidu is serving the DeepSeek V4 Flash 0731 model at extremely fast speeds and very low prices. The model is a sparse mixture-of-experts with 13B active parameters out of 284B total, designed for coding, reasoning, and agent workflows. OpenRouter lists the price at $0.0786 per 1M input tokens with 1M context support.

Original post →

More from Infra

Infra channel →