Kimi-K3 Tops LisanBench as Strongest Open-Weight Model
Kimi-K3 has become the strongest open-weight model on the LisanBench benchmark, outperforming Gemini 3. It ranks 5th in standard metrics and 4th in difficulty-weighted metrics, though its operating cost remains high.
2026-07-18 ~ 2026-07-18 · 2 related posts
- Episode 1: Kimi K3 Matches Top Models in Agentic Coding, but Real Cost Comes Under Fire(2026-07-18, 6 posts)
- Episode 2: Kimi-K3 Tops LisanBench as Strongest Open-Weight Model(2026-07-18, 2 posts)
- Episode 3: Kimi K3 Beats GPT-5.5 in Game Generation Test(2026-07-18, 2 posts)
- Episode 4: Kimi K3 Stuns with Coding and 3D Reasoning, Beating SOTA Models(2026-07-18, 5 posts)
- Episode 5: Kimi K3 Evaluations Show Polarized Results and Harness Sensitivity(2026-07-18, 5 posts)
- Episode 6: Kimi K3 Architecture Preview: Native Innovation and Attention Residuals(2026-07-19, 3 posts)
- Episode 7: Kimi-K3 Preliminary ECI Score Surpasses Top Models(2026-07-19, 4 posts)
- Episode 8: Kimi K3 Leads Harvey Legal Benchmark(2026-07-19, 4 posts)
- Episode 9: Moonshot Releases 2.8T Open-Weights Model Kimi K3(2026-07-19, 14 posts)
- Episode 10: Kimi K3 Open-Weight Model Ranks Top 3 Globally, Gap to Closed-Source Narrows to 4 Points(2026-07-28, 5 posts)
- Analyzing Kimi-K3 on the LisanBench Evaluation — scaling01 · 2026-07-18
- Kimi-K3 Review: Top Open-Source but Costly — scaling01 · 2026-07-18