K2 Horizon's data recipe: 20T tokens per model, 10T synthetic, 17% reasoning traces

rohanpaul_ai · x · 2026-09-11

Details of K2 Horizon's training data:

Related event: IFM Open-Sources K2 Horizon: Six Models from 0.9B to 375B with Parallel Decoding and Benchmark Cheating Audit(7 posts)→

Original post →

More from Models

Models channel →