AI safety testing is broken on all three fronts: labs, paid auditors and nonprofits all face warped incentives

joshua_saxe · x · 2026-09-11

Joshua Saxe argues the AI industry's safety testing regime is broken: labs self-test and report non-peer-reviewably, for-profit third parties depend on lab access for their valuations, and labs cherry-pick nonprofits whose reputations hinge on that access. Reacting to ccatalini's point that credible evidence of the agent capability-safety gap sits inside the labs, he calls this a prime place for policymakers to step in.

Original post →

More from Safety

Safety channel →