MiMo V2.5 Pro beats DeepSeek V4 Pro in a World Cup prediction test
PreciousSeige · reddit · 2026-07-28
A user compares MiMo V2.5 Pro and DeepSeek V4 Pro on soccer prediction tasks during the World Cup, using SportEval to track how each model performed.
Their takeaway is that MiMo V2.5 Pro ended up producing more accurate predictions in practice. The author says MiMo was better at updating judgments as new information arrived and at combining multiple factors into a coherent final call, while DeepSeek felt stronger on broad knowledge and benchmark-style performance. The screenshot also shows a leaderboard where GPT 5.5 sits above GLM 5.2, MiMo V2.5 Pro, Claude Opus 4.8, Kimi K2.6, Qwen3.7 Max, Seed2.1 Pro, DeepSeek V4 Pro, and Gemini 3.5 Flash.
More from Models
- Kimi K3 finds 16 new vulnerabilities and beats GLM-5.2 on an exploit benchmark — zephyr_z9 · 2026-07-28
- Kimi K3 weight shard appears as `model-00001-of-000096.safetensors` — ricklamers · 2026-07-28
- Microsoft launches MAI-Cyber-1-Flash and MDASH, claiming top CyberGym results at half the cost — satyanadella · 2026-07-28
- Claude Opus 5’s migration guide quietly changes years of prompting advice — AlexKim · 2026-07-28
- Post says an attention-heavy model has 104B active parameters — zephyr_z9 · 2026-07-28
- Kimi K3 hits 460 tokens/s on Modal as 5.6 Sol faces launch-week pressure — brandon_galang · 2026-07-28