MiMo near-SOTA on DeepSWE with just ~$2.6M RL run: will data cost more than training?

my_cat_can_code · x · 2026-09-21

mycatcancode raises a structural question about training economics: might data costs soon exceed compute costs for model training?

The evidence: MiMo achieved near-SOTA results on DeepSWE with an RL run costing only $2.6M. As RL training gets dramatically cheaper, acquiring and using high-quality data could become the dominant expense.

The post frames this as a potential shift in the training cost structure — compute getting cheaper while data gets more expensive.

Original post →

More from Infra

Infra channel →