Kimi K3 Scores Nearly Twice as High as Claude on Complex Legal Benchmarks

togethercompute · x · 2026-08-06

A tweet highlights that Kimi K3 scored nearly twice as high as Claude Fable 5 on Harvey LAB-AA’s hard autonomous legal tasks, demonstrating the model's strong capabilities in specialized vertical domains.

Original post →

More from Models

Models channel →