DeepSeek-V4-Flash-Vision-Exp Open Sourced, Topping Multimodal Benchmarks
智东西 · wechat · 2026-08-31
DeepSeek has open-sourced the experimental multimodal model DeepSeek-V4-Flash-Vision-Exp, integrating vision capabilities into the V4-Flash architecture. The model outperforms Opus-4.8 in 5 out of 7 text agent tasks and achieves a top score of 59.3% on the DeepSWE software engineering benchmark. In multimodal evaluations, it leads on ZeroBench and Agents'LastExam. However, it lags behind Opus-4.8 by about 10% on complex data science tasks like NL2Repo and DSBench-Hard. The DeepSeek-V4 series has seen 6 updates in the past month.
More from Models
- Google Releases Gemini Omni 1.1 Flash, Updating Its Fast Multimodal Model for Developers — thione · 2026-08-31
- DeepSeek launches low-cost vision model; Anthropic previews hardware control protocol for agents — thione · 2026-08-31
- Qwen and GLM release new MoE models focusing on low cost and high performance — thione · 2026-08-31
- Zhipu Releases Open-Weight GLM-5.3-Flash, a 320B MoE Model for Low-Cost Frontier Coding — thione · 2026-08-31
- DeepSeek releases V4-Flash-Vision-Exp, an open vision model with strong coding benchmarks — zainhas · 2026-08-31
- François Chollet explains ARC-3 evaluation on Kaggle — fchollet · 2026-08-31