DeepSeek releases V4.1-Flash, a multimodal API model with faster inference and lower prices
thione · x · 2026-09-15
- DeepSeek released V4.1-Flash, a multimodal API model promising faster inference, higher throughput, and lower prices, continuing its value-for-money approach.
- Exact pricing and benchmark details are in the official release.
- This closes the recap thread, in which the author aggregates recent AI news and promotes a weekly newsletter.
Related event: DeepSeek Releases V4.1-Flash with 890 Bytes/Token KV Cache(6 posts)→
More from Models
- Grok trains on user data by default; business plans can opt out — carlosdponx · 2026-09-15
- Bolt Forge launches free until Oct 14 with GLM, DeepSeek and Kimi plus up to 50x more usage — HeyAmit_ · 2026-09-15
- GPT-6 Astra tested on robot control: impressive on simple tasks, limited dexterity — DJiafei · 2026-09-15
- SOTA Inference Is Nearly Free for Consumers, So the Local-Model Trend May Reverse — mobileraj · 2026-09-15
- Cursor user switches to Claude Code, burns through quota by day 3 — jdluk87 · 2026-09-15
- Claude Max and Codex tiers are creating a computing power gap that locks out $20-budget newcomers — IndraVahan · 2026-09-15