Xiaomi MiMo-V2.6-Pro fixes real bugs at $0.86, carving out a strong Pareto frontier
PawelHuryn · x · 2026-09-22
After Xiaomi released MiMo-V2.6-Pro, Pawel Huryn blind-tested it on real work: 2 repos, 105 real bugs, with blind judges scoring (median, n≥3):
- Muse Spark 1.3 (max): 32.2 for $18.11
- GPT-5.6 Luna (max): 31.5 for $2.82
- Grok 4.7 (xhigh): 28.8 for $22.89
- MiMo-V2.6-Pro (default): 22.7 for just $0.86
- DeepSeek V4.1 Flash (max): 21.7 for $0.78
- Gemini 3.8 Flash (high): 18 for $11.03
Verdict: not the top scorer, but its price makes for a strong Pareto frontier — a viable default if subsidized subscriptions disappeared.
Related event: Bug Hunt Bench Tests Models on 105 Real Bugs, Cost Gap Nears 200x(3 posts)→
More from Models
- MiMo-V2.6 Flash and Pro Both Pass a 100-Step Clinical Agent Workflow That V2.5-Pro Failed — MaziyarPanahi · 2026-09-22
- GPT-6 Astra tops Claude Opus 5 in Ramp business spending share, OpenAI overtakes Anthropic — firstadopter · 2026-09-22
- 86 physics questions, 100 runs each: model hits 87.6% accuracy but understates confidence — Ok-Challenge-7810 · 2026-09-22
- DeepSeek reportedly bets on Huawei chips to train next-gen models; Liang says it 'has to work' — kimmonismus · 2026-09-22
- Xiaomi's MiMo-V2.6-Pro tops open models on $2.62M RL; Anthropic alleges Claude distillation — The Decoder · 2026-09-22
- Tencent finally opens WeChat interface, unlocking 100GB+ chat data processing — Xianbao_QIAN · 2026-09-22