Artificial Analysis Launches Harvey Legal Agent Benchmark
ArtificialAnlys · x · 2026-07-08
Artificial Analysis introduced Harvey LAB-AA (Legal Agent Benchmark), their implementation of Harvey's new agentic legal benchmark. It evaluates language models on real-world legal work across 24 practice areas.
The benchmark tests models on 120 private legal tasks built by the Harvey team, covering M&A, capital markets, tax, litigation, and bankruptcy. Models must produce specified legal deliverables, with each task scored on a binary scale.
Related event: Artificial Analysis Releases Harvey Legal Agent Benchmark Results(8 posts)→
More from Research
- Structural ensembles beat single predictions in TCR:pMHC generalization study — quaidmorris · 2026-07-22
- Structural ensembles, not single predictions, drive robust TCR:pMHC generalization — quaidmorris · 2026-07-22
- A 3D ray plot shows how hard this Jacobian counterexample is to read — moultano · 2026-07-22
- LLM leaderboards are now often measuring the harness too, Gary Marcus warns — GaryMarcus · 2026-07-22
- New paper defines self-state attacks, showing OS defenses leave four agent-memory cases indistinguishable — Justgototheeffinmoon · 2026-07-22
- Krea 2 users recommend a two-pass Clownshark sampler setup for sharper image details — listopalafoto · 2026-07-22