MiMo-V2.6 billed as largest single open-source RL run, harder than R1

teortaxesTex · x · 2026-09-22

Luo Fuli, who leads Xiaomi's MiMo team, introduced MiMo-V2.6, calling it likely the largest single RL run by compute ever undertaken by an open-source model team—keeping a dozens-strong team focused on scaling up RL despite scarce compute.

The recipe: mid-training to build potential, heavy RL to unlock it, resulting in what she claims is the #1 open-source model; she recommends the technical report as a future Agent RL classic.

She added that its research innovations and engineering challenges surpass DeepSeek R1, which she was partly involved in.

Related event: Xiaomi Open-Sources MiMo-V2.6; Pro Tops Open-Weight Intelligence Index(20 posts)→

Original post →

More from Models

Models channel →