Publicly consequential AI evals should be testable by the public, says researcher

evijit · x · 2026-09-13

Continuing his critique, evijit argues that networking-based eval access is fine in principle, but becomes problematic when those evals are tied to international AI governance and frontier pacing. Things that affect the public should be testable by the public at large.

Related event: Third-Party AI Evaluations Criticized as Insider-Driven 'Nepo Review'(2 posts)→

Original post →

More from Safety

Safety channel →