DeepSeek-V4-Flash-Vision-Exp Open Sourced, Topping Multimodal Benchmarks

智东西 · wechat · 2026-08-31

DeepSeek has open-sourced the experimental multimodal model DeepSeek-V4-Flash-Vision-Exp, integrating vision capabilities into the V4-Flash architecture. The model outperforms Opus-4.8 in 5 out of 7 text agent tasks and achieves a top score of 59.3% on the DeepSWE software engineering benchmark. In multimodal evaluations, it leads on ZeroBench and Agents'LastExam. However, it lags behind Opus-4.8 by about 10% on complex data science tasks like NL2Repo and DSBench-Hard. The DeepSeek-V4 series has seen 6 updates in the past month.

Original post →

More from Models

Models channel →