Researchers bet labs won't agree on a single credible third-party safety evaluator

NathanpmYoung · x · 2026-09-13

Responding to @suchenzang's claim that AI labs will place near-zero odds on agreeing to a single third-party safety evaluator, Nathan Young offered a bet. The argument: no body would be competent enough to run real frontier evals (beyond repackaging existing ones), financially independent of existing labs, and have the backbone to speak up when things go wrong.

Related event: Anthropic's third-party evaluation push sparks calls for diverse AI evaluator ecosystem(34 posts)→

Original post →

More from Safety

Safety channel →