DeepSeek reportedly admits V4-Pro pretraining misstep; V4.1-Flash due Sept 10
teortaxesTex · x · 2026-09-09
Community sources say DeepSeek has conceded a pretraining mistake on V4-Pro, which will be replaced by V4.1-Flash — officially releasing September 10th. Commenter teortaxesTex hopes to learn what went wrong, and optimistically reads V4.1-Flash as the intended compression of the V4 design that may actually scale this time. Unconfirmed by DeepSeek.
More from Models
- Nvidia's Jensen Huang: closed models are cheaper, open models give you control — rohanpaul_ai · 2026-09-09
- What to watch after Meta's persistent VM with a built-in gatekeeper: Sentinel, Muse, and rivals — eyishazyer · 2026-09-09
- Analyst speculates OpenAI achieved multi-agent swarm breakthrough, with agents self-organizing since May — teortaxesTex · 2026-09-09
- GPT-6 called the most interesting model release in years, thanks to HF and German forum leaks — xeophon · 2026-09-09
- Anthropic Max users report Opus quietly excluded from 'all models' usage bar — wyongriver · 2026-09-09
- DeepSeek v4.1 Generates 100 Distinct HTML Artworks in One Go, Blogger Says It Beats GPT-6 Astra — teortaxesTex · 2026-09-09