Kimi K3 Weights Released: 93 Layers, Latent MoE Replaces Standard MoE
hugobowne · x · 2026-07-29
Kimi K3 weights dropped yesterday. Sebastian Raschka quickly analyzed the architecture:
- 93 layers (vs. 27 in Kimi Linear)
- Latent MoE replaces standard MoE
- Raschka's summary: "It's basically Kimi Linear, bigger, with latent MoE."
Full conversation at the video link.
More from Models
- LiquidAI’s 230M LFM2.5 encoder trends on Hugging Face — LiquidAI · 2026-07-29
- Moonshot’s Kimi K3 is a 2.8T open-weight MoE model with 1M-token context — alex_verem · 2026-07-29
- Kimi K3 Tech Report: How Moonshot Achieved 2.5x Compute Efficiency — alex_verem · 2026-07-29
- Experienced web developer says Claude Opus 4.8 and 5 are now unusable for chat — FuzzyHead455 · 2026-07-29
- Reddit user says paid Gemini Pro access is still routing to other models — IAmMonke2 · 2026-07-29
- A fake Claude “model welfare” leak turns into an AI-community meme — repligate · 2026-07-29