Leaked: DeepSeek V4.1 Flash open weights, claims 86x cheaper and 420+ tok/s
ns123abc · x · 2026-09-10
An unverified post claims DeepSeek dropped V4.1 Flash with fully open weights, beating GPT 5.6 Sol, Opus 5 and every Chinese model on coding and cybersecurity at 86x lower cost per million tokens, running 420-507 tok/s (faster than Gemini 3.8 Flash), described as the "smallest model in our new architecture family." None of this is officially confirmed — treat as a rumor pending DeepSeek's announcement.
More from Models
- Users Petition OpenAI for $400-$600 Heavy Builder Tier as $200 Plan Runs Dry in 48 Hours — dragonwarrior_1 · 2026-09-10
- Leak: 'SpaceXAI' working to bring Grok Bots into XChat for in-conversation tagging — nima_owji · 2026-09-10
- Chinese model's 74.2 score under fire: best of 8 eval variants, maxed thinking budget, ~2.5x cost — teortaxesTex · 2026-09-10
- Unitree fully open-sources UnifoLM-WLA-1.0, a 6B humanoid robot foundation model — teortaxesTex · 2026-09-10
- DeepSeek's new release shows ChatGPT fingerprints, token efficiency set to jump — teortaxesTex · 2026-09-10
- Analyst: DeepSeek's latest change is a big win for token efficiency, moving toward OpenAI's regime — teortaxesTex · 2026-09-10