DeepSeek V4.1 Flash launches: 552B MoE backbone, 1M-token context
tiguidoio · reddit · 2026-09-10
A Reddit user reports the release of DeepSeek's new model V4.1 Flash:
- Multimodal Mixture-of-Experts (MoE) architecture
- 552B backbone parameters
- Supports up to 1 million tokens of context
The poster quips it's "market crash as a service," poking fun at the competitive impact on rivals. More benchmarks and licensing details are in the official HF repo.
More from Models
- Leaked brief: DeepSeek V4.1 Flash at 552B params, foldable iPhone at $1,999, ChatGPT voice limits raised — testingcatalog · 2026-09-10
- CoT similarity test suggests Qwen3.8 may have been trained on GPT 5.5 reasoning traces — Chromix_ · 2026-09-10
- DeepSeek accused of 'pretending linear attention doesn't exist' in new architecture — teortaxesTex · 2026-09-10
- Hands-on with GLM 5.3 Flash: great at coding and research, poor at trading and ideation — ManagementNo5153 · 2026-09-10
- Tell an agent it has 1M token budget and it reasons 3-5x longer: a metacognition experiment — paraschopra · 2026-09-10
- DeepSeek's new open model beats GLM 5.3 and Kimi K3 at 4-10x lower price — deedydas · 2026-09-10