DeepSeek V4.1 Flash OCR Benchmarked: 260 tok/s, Competitive Accuracy, Low Cost
solyarisoftware · x · 2026-09-10
NFTChen's test shows DeepSeek V4.1 Flash multimodal OCR achieves 6 errors + 1 miss on 270-character running script, tying GLM 5.3 Flash and Gemini 3.1 Pro; Kimi2.6 scores perfect, Qwen3.8-Max 5 errors. Speed: native multimodal stable at 260 tok/s (old version 100-150), single-image OCR under 3s with thinking off, others 15-50s. Cost-wise, Kimi/Qwen/GLM are accurate but expensive; DeepSeek V4.1 Flash after price cut balances accuracy and cost, ideal for large-scale ancient text OCR. Also, DeepSeek V4 Pro was replaced by V4.1 Flash after just 28 days, with Pro requests now routed to Flash pricing.
More from Models
- ChatGPT Voice gets usage caps: 3h for Plus, 15h for Pro $100, $200 stays unlimited — testingcatalog · 2026-09-10
- VDiff-Bench: 1,756-question benchmark shows frontier models fail at spot-the-difference — yixin_wan_ · 2026-09-10
- OpenAI team points users to official usage limit update details — athyuttamre · 2026-09-10
- MiniCPM5-2B: 2B-parameter open model runs agents offline on 2GB RAM, tops sub-4B open models — solyarisoftware · 2026-09-10
- ChatGPT Voice gains GPT-5.6 Sol and GPT-6 Astra, Plus/Pro usage limits raised — athyuttamre · 2026-09-10
- Rumor: DeepSeek V4.1 Flash to launch ~Sept 10, V4 Pro requests rerouted at Flash pricing — solyarisoftware · 2026-09-10