Xiaomi open-sources MiMo-V2.6, scaling RL to 1,568 samples and 3.7B tokens per step

XiaomiMiMo · hf · 2026-10-09

Xiaomi's MiMo team released MiMo-V2.6, an omni-modal model family that treats scaled RL compute as the central path to self-improvement, open-sourcing training dynamics, RL environments, and the RL framework.

Key points:

Original post →

More from Models

Models channel →