DeepSeek V4 Flash Hits OpenRouter with 1M Context at $0.14 Input Pricing
michellechen · x · 2026-08-01
OpenRouter has officially listed the DeepSeek V4 Flash 0731 model. According to the platform, this model utilizes a sparse mixture-of-experts (MoE) architecture with 13B active parameters out of 284B total, optimized for coding, reasoning, and agent workflows.
Pricing is set at $0.14 per 1M input tokens and $0.28 per 1M output tokens, featuring a massive 1M context window. OpenRouter offers various routing modes, including Balanced, Nitro (fastest), and Exacto (highest tool-calling accuracy).
More from Models
- Fable's AI Safety Filter Constantly Triggers on Benign Content — dreamwieber · 2026-08-01
- DeepSeek-V4-Flash Inference Blocked: vLLM Lacks Support for New confidence_head — teortaxesTex · 2026-08-01
- Without Open-Weight AI, Closed Models Could Cost $2,000/Month, Says KOL — iamaliveix · 2026-08-01
- APEX-Accounting Benchmark: 58% Tasks Unsolved, Claude Fable 5 Takes the Lead — EdwardSun0909 · 2026-08-01
- Hands-on with GPT-5.6 Luna: Matches Sol in Knowledge Work at a Fraction of the Cost — BenBajarin · 2026-08-01
- OpenAI Slashes Prices: GPT-5.6 Terra and Luna Now 50% Off — LeTanLoc98 · 2026-08-01