Kimi K3 pairs a 2.8T MoE with 1M context and multi-teacher post-training

novasarc01 · x · 2026-07-28

Kimi K3 is a 2.8T MoE model with native vision and 1M-token context

The post summarizes the Kimi K3 technical report and its training recipe.

Related event: Kimi K3 Report: Trillion-Parameter MoE and Post-Training Innovations(2 posts)→

Original post →

More from coding & agent

coding & agent channel →