Ex-OpenAI researcher: odds near zero labs agree on a real third-party safety evaluator

basedjensen · x · 2026-09-13

Responding to a proposal for a unified third-party frontier safety evaluator, ex-OpenAI researcher Steven Adler puts near-zero probability on labs agreeing to one, citing three barriers: few orgs can run genuine frontier evals (vs repackaging existing ones), financial ties to existing labs are hard to avoid, and few evaluators would have the backbone to speak up when things go wrong. A commenter adds that nearly everyone in the space has ties to EA, compounding independence concerns.

Related event: Ex-OpenAI Researchers: Unified Third-Party AI Safety Evaluators Nearly Impossible(3 posts)→

Original post →

More from Safety

Safety channel →