MiMo near-SOTA on DeepSWE with just ~$2.6M RL run: will data cost more than training?
my_cat_can_code · x · 2026-09-21
mycatcancode raises a structural question about training economics: might data costs soon exceed compute costs for model training?
The evidence: MiMo achieved near-SOTA results on DeepSWE with an RL run costing only $2.6M. As RL training gets dramatically cheaper, acquiring and using high-quality data could become the dominant expense.
The post frames this as a potential shift in the training cost structure — compute getting cheaper while data gets more expensive.
More from Infra
- How do you test LLM provider failure in production? A Reddit discussion — Rama_Surasani_ · 2026-09-21
- Run Flux 2 Dev (30B) locally with block offloading, mix models for T2I and editing — Altruistic_Heat_9531 · 2026-09-21
- Jensen Huang stands by $3-4T AI infrastructure market forecast — emmanuelvivier · 2026-09-21
- Free Zoom meetup: disaggregated speculative decoding on d-Matrix chips plus inference engine tuning — cfregly · 2026-09-21
- Program-as-Weights: 0.6B model matches Qwen3-32B prompting with 1/50th memory, runs locally — yuntiandeng · 2026-09-21
- Personal AI Agents Are the Biggest Driver of the Sudden NAND Demand Surge — zephyr_z9 · 2026-09-21