Musk proposes AI safety peer review: Anthropic tests OpenAI, everyone breaks everyone's models pre-release

elonmusk · x · 2026-09-15

At the All-In Summit, Elon Musk offered a practical framing of the AI safety problem: OpenAI and Anthropic are close enough in capability that neither can slow down without handing the lead to the other. He noted Anthropic puts more care into safety than OpenAI, yet even Anthropic insiders publicly worry about their own models.

His proposed fix instead of voluntary restraint: cross-testing. Anthropic runs its safety test harness on OpenAI models, OpenAI tests Anthropic, SpaceXAI tests both, and leading Chinese AI companies join the same system — everyone tries to break everyone else's models before release. The full discussion is in the latest All-In episode, alongside Starship, space data centers, and Terafab topics.

Related event: Musk Proposes AI Labs Cross-Audit Each Other's Safety Tests(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →