Artificial Analysis and Harvey add hallucination gating to Legal Agent Benchmark (LAB-AA v1.1)

ArtificialAnlys · x · 2026-10-09

Artificial Analysis and Harvey announced LAB-AA v1.1, updating the Legal Agent Benchmark scoring to add a hallucination check. The new headline metric, Hallucination-Gated All-Pass Rate, only credits a task when deliverables satisfy every rubric criterion and contain no material misstatements.

Related event: Hallucination gating reshuffles legal agent benchmark; Grok 4.7 takes the lead(8 posts)→

Original post →

More from Models

Models channel →