DeepSeek v4.1 Flash preview spotted: ~300 tok/s, fires lots of subagents, no vision yet

kevinkern · x · 2026-09-10

Developer Kevin Kern tested an upcoming DeepSeek v4.1 Flash preview endpoint, reporting inference at roughly 300 tok/s and heavy parallel subagent usage on his workloads.

The preview still lacks vision; he hopes the official release will be multimodal.

Related event: DeepSeek v4.1 flash preview spotted running at ~300 tok/s, vision still missing(2 posts)→

Original post →

More from Models

Models channel →