Moonshot releases Kimi K3 report, claiming 2.5× scaling-efficiency gain over K2
cedric_chee · x · 2026-07-27
Moonshot has released the Kimi K3 technical report, claiming a 2.5× gain in scaling efficiency over Kimi K2.
The report attributes the improvement to a combination of architecture, data, and training changes. The attached table highlights the main structural differences, including a larger MoE stack, more activated parameters, more experts per token, a longer training context, and a shift to a hybrid KDA–MLA attention design.
Related event: Moonshot releases Kimi K3 open weights amid license debate(155 posts)→
More from Models
- theo builds his own visualizer for today's agent models, showing how cheap Luna really is — ivan_bezdomny · 2026-09-23
- Why ChatGPT Still Wins: One User's Split Between Muse, Claude and Codex — mobileraj · 2026-09-23
- Muse reportedly offers 4B tokens/week for ~$100/month, sparking industry price-disruption talk — NewYak4281 · 2026-09-23
- GPT-6 Sol and Luna appear in OpenAI docs, alongside guidance on reasoning effort — cedric_chee · 2026-09-23
- GPT-6 tested on LIBERO robot task: turns on stove, fails to grasp moka pot — YuXiang_IRVL · 2026-09-23
- Ternary Bonsai 2 27B: 5.9GB weights retain ~95% of full-precision reasoning — cephaloform · 2026-09-23