DeepSeek open-sources V4 series first multimodal model: Flash-Vision-Exp

赛博禅心 · wechat · 2026-08-31

DeepSeek has open-sourced DeepSeek-V4-Flash-Vision-Exp, the first experimental multimodal model in the V4 series. Built on the DeepSeek-V4-Flash architecture, it gains visual understanding capabilities through introduced vision modules and continuous training. Compared to V4-Flash-0731, it shows significant improvements in multimodal Agent capabilities while maintaining comparable performance in pure text Agent tasks.

Related event: DeepSeek Quietly Open-Sources V4-Flash-Vision-Exp, Its First Multimodal Model(12 posts)→

Original post →

More from Models

Models channel →