DeepSeek V4.1 Flash drops: 552B MoE activating just 8B, beating GPT-5.6 Sol at $0.14/M

ChrisGPT · x · 2026-09-10

DeepSeek has released V4.1 Flash, a 552B-parameter MoE model that activates only 8B parameters for input and 16B for output. It beats GPT-5.6 Sol and Opus 5 on several agentic/coding benchmarks: CyberGym 88.1 vs 84.5, Automation-Bench 54.8 vs 45.8, and DeepSWE 74.2 vs 73.0. Off-peak API pricing is roughly $0.14/M input and $0.56/M output tokens.

Related event: DeepSeek Unveils Open-Source V4.1-Flash MoE Model(28 posts)→

Original post →

More from Models

Models channel →