DeepSeek unveils V4.1-Flash: smallest model in new architecture family with native vision
zephyr_z9 · x · 2026-09-10
DeepSeek officially introduced V4.1-Flash, the smallest model in its new architecture family, with native visual understanding, designed for greater capability, faster inference and higher throughput as a stepping stone to larger models. Retweeter Dorialexander quipped "model is data." The model offers 552B backbone params (16B active), 1M context, and day-0 support in SGLang.
More from Models
- Leak: 'SpaceXAI' working to bring Grok Bots into XChat for in-conversation tagging — nima_owji · 2026-09-10
- Chinese model's 74.2 score under fire: best of 8 eval variants, maxed thinking budget, ~2.5x cost — teortaxesTex · 2026-09-10
- Unitree fully open-sources UnifoLM-WLA-1.0, a 6B humanoid robot foundation model — teortaxesTex · 2026-09-10
- DeepSeek's new release shows ChatGPT fingerprints, token efficiency set to jump — teortaxesTex · 2026-09-10
- Analyst: DeepSeek's latest change is a big win for token efficiency, moving toward OpenAI's regime — teortaxesTex · 2026-09-10
- DeepSeek V4.1 Flash forward pass walkthrough shows new efficiency tricks — Thom_Wolf · 2026-09-10