PNAS Paper Explores Institutional Design of Legal AI Benchmarking
chrmanning · x · 2026-07-22
Neel Guha and co-authors have published a new piece in PNAS examining legal AI benchmarking from an institutional perspective.
The article argues that while much has been written on the technical aspects of benchmarking—such as picking metrics and building datasets—relatively little attention has been paid to institutional design. Although focused on legal AI, the research offers two broader ideas highly relevant to general AI governance conversations.
More from Safety
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- California creates standards for independent AI auditors to verify lab safety testing — VraserX · 2026-09-11
- a16z podcast: why 2-3 person startups are absent from policy debates — a16z Podcast · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- Class action accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers — The Decoder · 2026-09-11
- MD shows buying lab media requires background checks, calling AI bioweapon doom scenarios implausible — Ghost_Pilot_MD · 2026-09-11