DeepSeek Unveils V4.1-Flash, a 552B MoE Redesigned for Agents

DeepSeek released V4.1-Flash, its smallest new-architecture model: a 552B-parameter MoE with only 8B active, featuring an asymmetric Causal Encoder-Decoder design and KV cache compressed to 890 bytes per token for agent workloads.

2026-09-21 ~ 2026-09-21 · 3 related posts