DeepSeek’s next model should add continual learning before chasing bigger gains
teortaxesTex · x · 2026-07-23
The thread argues that DeepSeek’s next generation should prioritize continual learning. It points to the V4 paper and hiring signals as evidence that the team is moving toward that direction.
A reply adds that V4 should also have native multimodality, while another commenter suggests V4 is already at the largest scale DeepSeek can train right now. The implication is that the company may need to keep improving cost/performance before attempting the bigger leap.
Related event: Leaked DeepSeek Recording Reveals AGI and Open Source Strategy(37 posts)→
More from Models
- Simon Willison says loops are becoming obsolete as models handle long tasks natively — teropa · 2026-07-23
- A rumored Claude Opus 5, OpenAI Codex voice agents, and 750 tok/s from Cerebras — imjustnewatai · 2026-07-23
- A BrowseComp chart puts Kimi K3 near the top on score while keeping cost low — iamfakhrealam · 2026-07-23
- Grok Build reportedly beats Codex and Claude Code on browser tasks — elonmusk · 2026-07-23
- Reddit users ask whether Hunyuan Image 3.0 Instruct is worth 170GB VRAM — dtdisapointingresult · 2026-07-23
- Real Task Cost Across GPT, Claude, Gemini, Kimi: 10.6x Spread Despite Only 2x Price Difference, Hidden Reasoning Tokens Blamed — pixelo2323 · 2026-07-23