Kimi K3 Leads Legal Benchmarks
ZainHasan6 · x · 2026-07-19
Kimi K3 has shown a significant lead on Harvey's legal task benchmarks. According to the charts shared, Kimi K3 achieved an **all-pass rate** of **26.7%** on Harvey LAB-AA, nearly double the **14.2%** scored by Claude Fable 5. This benchmark measures autonomous, real-world legal work rather than simple Q&A.
Related event: Kimi K3 Leads Harvey Legal Benchmark(4 posts)→
More from Models
- Kimi post pairs a compute anecdote with a subjective top-5 model ranking — vista8 · 2026-07-21
- DavidAU collaboration reportedly improves Qwen 3.6 27B on long-context agent tasks — My_Unbiased_Opinion · 2026-07-21
- Smaller language models stay terse while larger ones explain the physics — NickPassig · 2026-07-21
- Sources and Lean proof links for the IMO 2026 model benchmark — deedydas · 2026-07-21
- Claude Fable, GPT-5.6 Sol, Kimi K3 all score 42/42 on IMO 2026 — deedydas · 2026-07-21
- Kimi K3 tested against Claude Code in the same Pi harness workflow — tomcrawshaw01 · 2026-07-21