Who evaluates the independent evaluators? Philosopher flags oversight gap in AI safety reviews

connoraxiotes · x · 2026-09-13

Philosopher Jeff Sebo raises a recursive problem for AI safety: as frontier labs rely on "independent evaluators," who audits the auditors?

Even without outright collusion, institutional relationships — funding and partnership ties — can compromise review integrity. His proposed remedy is meaningful democratic oversight: public accountability for how evaluators are selected and supervised, rather than relying on industry self-regulation alone.

Original post →

More from AGI Musings

AGI Musings channel →