Peter Gostev debunks model sparsity leak: Kimi 26:1, DeepSeek 32:1, 1.2T active params implausible

inductionheads · x · 2026-09-07

Peter Gostev pushed back on a viral leak claiming an OpenAI model with 1.2T active parameters, running the MoE sparsity math: Kimi k3 is 2.8T total with 104B active (26:1), and DeepSeek v4 is sparser still at 1.4T total to 49B active (32:1). Given the industry trend toward even higher sparsity, he argues there's no plausible way OpenAI would sit at 5:1 or 10:1 — making the 1.2T claim clearly fabricated.

Original post →

More from Models

Models channel →