DeepSeek’s next model should add continual learning before chasing bigger gains

teortaxesTex · x · 2026-07-23

The thread argues that DeepSeek’s next generation should prioritize continual learning. It points to the V4 paper and hiring signals as evidence that the team is moving toward that direction.

A reply adds that V4 should also have native multimodality, while another commenter suggests V4 is already at the largest scale DeepSeek can train right now. The implication is that the company may need to keep improving cost/performance before attempting the bigger leap.

Related event: Leaked DeepSeek Recording Reveals AGI and Open Source Strategy(37 posts)→

Original post →

More from Models

Models channel →