Xiaomi MiMo Finishes RL Training, DeepSWE Score Jumps to 72.57
Xiaomi's MiMo completed reinforcement learning training, boosting its DeepSWE score from 58.41 to 72.57—near the 74 record—while MiMo-V2.6 emerged as the latest open-source model to publicly document its RL training.
2026-09-21 ~ 2026-09-23 · 2 related posts
- Episode 1: Xiaomi MiMo Finishes RL Training, DeepSWE Score Jumps to 72.57(2026-09-21, 2 posts)
- Episode 2: Xiaomi Open-Sources MiMo-V2.6, Topping Open-Weights Leaderboard with Livestreamed RL Training(2026-09-22, 39 posts)
- Episode 3: Xiaomi's MiMo-V2.6 Pro tested: big gains in coding, 3D games and multimodal(2026-09-22, 2 posts)
- Xiaomi's MiMo Pro jumps from 58.41 to 72.57 on DeepSWE after RL runs — nrehiew_ · 2026-09-21
- MiMo-V2.6 streams RL run: DeepSWE score jumps from 58.41 to 72.57, nearing SOTA — teortaxesTex · 2026-09-23