DeepSeek V4 Specs and Pricing Leaked: Supports 1M Context
dotey · x · 2026-08-12
DeepSeek's API documentation page has revealed two suspected V4 series models: deepseek-v4-flash and deepseek-v4-pro.
- Core Specs: Both models support a massive 1M context length and up to 384K output tokens. They feature JSON Output, Tool Calls, and compatibility with both OpenAI and Anthropic API formats.
- Thinking Mode: The Pro version supports switching between thinking and non-thinking modes by default, while Flash also supports it.
- Pricing & Limits: Flash costs 1 RMB/M tokens for cache-miss input, and Pro costs 3 RMB. The docs note an upcoming significant price increase for the API service. Concurrency limits are set at 2500 for Flash and 500 for Pro.
Related event: Leaked Specs and Pricing for DeepSeek V4 Spark Community Buzz(4 posts)→
More from Models
- Dev Reflects on AI Coding: Models Make Basic Reasoning Errors, Hand-Coding Wins — jsuarez · 2026-08-13
- Grok Hits SOTA on Databricks' OfficeQA Pro V2 Evaluation — GavinSBaker · 2026-08-13
- Grok 4.6 hits Vercel AI Gateway with 500K token context window — soleio · 2026-08-13
- xAI Grants SuperGrok Users a Free Weekly Usage Reset — XFreeze · 2026-08-13
- Krea, FLUX, and MiniMax Open-Weights: Promises vs Reality — felixsanz · 2026-08-13
- Anthropic Slammed by Users for Extra 'Fast Mode' Fees on Claude Code — Neel_MynO · 2026-08-13