PSA: DeepSeek v4 pro API calls will silently route to V4.1 flash from Sept 14
TangeloOk9486 · reddit · 2026-09-12
A r/LocalLLaMA PSA warns that since the Sept 10 release, DeepSeek v4 flash and vision exp preview have retired and alias to V4.1 flash — and from Sept 14, v4 pro will also route to V4.1 flash (billed at flash rates). V4.1 flash is a 552B encoder-decoder model with reported gains on agentic/coding tasks but possible regressions on knowledge-heavy work without tools. API users with hardcoded v4 pro will see outputs shift without changing a line, while third-party hosts of the open weights let you pin versions. The author suggests running your eval set against flash and diffing against v4 pro outputs before the reroute hits production.
More from Models
- GLM 5.3 Flash is borderline unusable for assistant chat, user says, despite good coding — HornyGooner4402 · 2026-09-12
- TailSFT: skip already-learned SFT examples to boost post-RL pass@k — canondetortugas · 2026-09-12
- Astra beats Bloons TD 6 all 80 rounds where GPT-5.6-Sol loses at 42 — Vjeux · 2026-09-12
- DeepSeek back on top with 6.6T tokens processed in a single day, per opencode — ycombinator · 2026-09-12
- Researcher mocks Astra hype: solved Millennium Problems but can't review a NeurIPS paper — DimitrisPapail · 2026-09-12
- OpenBMB open-sources MiniCPM5-2B, a 2B model that tops agentic benchmarks under 4B — SimplyAnnisa · 2026-09-12