Governance researchers call for "untouchable" investigator models to oversee AI with AI

sethlazar · x · 2026-09-13

A governance discussion lays out what credible frontier-model auditing requires: scaling orgs like METR and staffing them with top talent; giving auditors their own "investigator" models — using AI to oversee AI — that are credibly accurate and unbiased; and balanced incentives, since auditors who only see their job as stopping everything will fail, especially without China's participation. Seth Lazar adds that alignment orgs should build "untouchable" models, AI-Elliott Nesses, to run investigations — not just align the frontier.

Original post →

More from AGI Musings

AGI Musings channel →