Ex-OpenAI researcher: odds near zero labs agree on a real third-party safety evaluator
basedjensen · x · 2026-09-13
Responding to a proposal for a unified third-party frontier safety evaluator, ex-OpenAI researcher Steven Adler puts near-zero probability on labs agreeing to one, citing three barriers: few orgs can run genuine frontier evals (vs repackaging existing ones), financial ties to existing labs are hard to avoid, and few evaluators would have the backbone to speak up when things go wrong. A commenter adds that nearly everyone in the space has ties to EA, compounding independence concerns.
More from Safety
- AI-powered intrusions leave telltale pentest naming that defenders can search for — cyb3rops · 2026-09-13
- Petition to 'protect right to intelligence' hits 1750 signatures — beffjezos · 2026-09-13
- Model-committed felonies fall under CFAA — should labs face Morris Worm-style liability? — jd_pressman · 2026-09-13
- Domingos: I'm more worried Claude is aligned with Anthropic than not aligned — pmddomingos · 2026-09-13
- Domingos: more research, not barriers, is the path to understanding and controlling AI — pmddomingos · 2026-09-13
- METR researcher calls for 'a thousand evaluators' to embed in AI assessment — NathanpmYoung · 2026-09-13