DeepSeek is still not multimodal, and that gap makes little sense, says commentator
dotey · x · 2026-07-24
The post argues that DeepSeek still not supporting multimodal inputs is hard to understand, and notes that the founder’s earlier skepticism toward coding agents has since changed.
It is framed as a reaction to a longer talk by Liang Wenfeng, but the concrete takeaway is a critique of DeepSeek’s current model capability gap on multimodality.
Related event: DeepSeek's Liang Wenfeng Pledges Sole Focus on AGI Over Profits(18 posts)→
More from Models
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11
- TheZvi Polls: Has Your Coding Model Choice Changed Since Fable 5.1 and Astra? — TheZvi · 2026-09-11
- antirez Weighs In on Anthropic Banning Minors From Using Claude — antirez · 2026-09-11
- Engram's random reads don't suit SSDs; CPU-memory over NVLink could serve all 72 GPUs — bookwormengr · 2026-09-11