DeepSeek claims V4.1-Flash beats V4-Pro, will route all V4-Pro API traffic to it from Sept 14
eyishazyer · x · 2026-09-11
DeepSeek claims its V4.1-Flash beats V4-Pro on performance, cost, speed, and total runtime — enough that all V4-Pro API traffic will be routed to Flash starting Sept 14. Per a notice relayed by TechNode, off-peak pricing for cache-hit input tokens reportedly drops to around RMB 0.02 per million tokens, though this hasn't yet appeared on DeepSeek's official pricing page.
More from Models
- ChatGPT monthly active users top 1.06 billion in August, fourth straight record month — FinanceYF5 · 2026-09-11
- PuzzleMask: Plain-Prose Attack Bypasses All 4 Tested LLM Gatekeepers at 100% — TechNadu · 2026-09-11
- OpenAI Codex may issue another usage reset this weekend, says Codex lead resets happen — umesh_ai · 2026-09-11
- OpenAI Reportedly Pointing Its Navier–Stokes Model at Riemann and P vs NP — 141_1337 · 2026-09-11
- Benchmark author says OpenRouter unreliably honors Meta Muse effort levels, EU payments broken — PawelHuryn · 2026-09-11
- User burns $200 of Codex credits in one agent turn — 4,700 of 5,000 credits, task unfinished — RileyRalmuto · 2026-09-11