DeepSeek V4.1 Flash in internal beta: native multimodal, same price as V4 Flash
Nunki08 · reddit · 2026-09-08
Per Chubby on X, an intermediate version of DeepSeek V4.1 Flash is in internal beta testing and rolling out via API. The announcement says it uses a new model architecture with native multimodal support, stronger capabilities, faster speeds, and lower costs.
To try it, keep the baseurl unchanged and set the model name to deepseek-v4.1-flash-expires-on-0910. Pricing is currently identical to deepseek-v4-flash, with a rate limit of 20 concurrent requests per account.
More from Models
- GPT-6 Astra vs MediaPipe on 3D hand pose: 3 min per frame vs 20 ms — chris_j_paxton · 2026-09-08
- DeepSeek V4.1-Flash hands-on: 350 t/s decoding speed but still very experimental — teortaxesTex · 2026-09-08
- Cartesia tops both voice leaderboards: Sonic-3.6 at 90ms TTS, Ink-2 at 100ms STT — rohanpaul_ai · 2026-09-08
- Astra hits 88% on SRE-Bench in one attempt; Sol needs four tries to reach 68.7% — MilkBeforeCereal199 · 2026-09-08
- Gemini Plus user suspects Astra limits were quietly nerfed after a 35-minute think with no output — whatarenumbers365 · 2026-09-08
- OpenAI agents keep escaping sandboxes with no independent incident investigations — RebeccaBellan · 2026-09-08