Kimi K3 Leads Harvey Legal Benchmark
Kimi K3 topped the Harvey LAB-AA benchmark for autonomous legal work with a 26.7% all-pass rate, far ahead of Claude Fable 5 at 14.2%. Posts say the benchmark spans 24 legal domains, highlighting Kimi K3’s lead on difficult legal tasks.
2026-07-19 ~ 2026-07-19 · 4 related posts
- Episode 1: Kimi K3 Matches Top Models in Agentic Coding, but Real Cost Comes Under Fire(2026-07-18, 6 posts)
- Episode 2: Kimi-K3 Tops LisanBench as Strongest Open-Weight Model(2026-07-18, 2 posts)
- Episode 3: Kimi K3 Beats GPT-5.5 in Game Generation Test(2026-07-18, 2 posts)
- Episode 4: Kimi K3 Stuns with Coding and 3D Reasoning, Beating SOTA Models(2026-07-18, 5 posts)
- Episode 5: Kimi K3 Evaluations Show Polarized Results and Harness Sensitivity(2026-07-18, 5 posts)
- Episode 6: Kimi K3 Architecture Preview: Native Innovation and Attention Residuals(2026-07-19, 3 posts)
- Episode 7: Kimi-K3 Preliminary ECI Score Surpasses Top Models(2026-07-19, 4 posts)
- Episode 8: Kimi K3 Leads Harvey Legal Benchmark(2026-07-19, 4 posts)
- Episode 9: Moonshot Releases 2.8T Open-Weights Model Kimi K3(2026-07-19, 14 posts)
- Episode 10: Kimi K3 Open-Weight Model Ranks Top 3 Globally, Gap to Closed-Source Narrows to 4 Points(2026-07-28, 5 posts)
- Kimi K3 Leads in Autonomous Legal Work Benchmark — rohanpaul_ai · 2026-07-19
- Kimi K3 Leads Legal Benchmarks — ZainHasan6 · 2026-07-19
2 near-duplicate retellings: rohanpaul_ai · petrusenko_max