Blogger corrects himself: the real surprise is MiMo-2.6, beating Grok 4.7 at much lower cost
kimmonismus · x · 2026-09-22
Blogger kimmonismus walked back his earlier pick after landing in SF: commenters convinced him the actual standout is MiMo-2.6. He claims it beats Grok 4.7 while costing much less, declaring "China is back" and saying he plans to test it himself. This is an early first impression, not a full benchmark.
More from Models
- Sentdex benchmarks openjev: 169ms on Dell GB10 vs 137ms on RTX 3090 — Sentdex · 2026-09-22
- Vals AI: Grok 4.7 drops to #24 on Vals Index, down 5 points from Grok 4.6 — zacharynado · 2026-09-22
- Grok 4.7 example: three-year financial analysis exposes currency-masked growth stall — ArtificialAnlys · 2026-09-22
- AA example: Grok 4.7 independently runs valuation chain and flags divergence from deal partner — ArtificialAnlys · 2026-09-22
- Grok 4.7 ranks just behind Anthropic's Opus 5 on AA-Briefcase at ~50% of the cost per task — ArtificialAnlys · 2026-09-22
- Replicating ExploitBench Would Cost ~$59.3M in API Fees, Security Researcher Estimates — OwariDa · 2026-09-22