Musk proposes AI safety peer review: Anthropic tests OpenAI, everyone breaks everyone's models pre-release
elonmusk · x · 2026-09-15
At the All-In Summit, Elon Musk offered a practical framing of the AI safety problem: OpenAI and Anthropic are close enough in capability that neither can slow down without handing the lead to the other. He noted Anthropic puts more care into safety than OpenAI, yet even Anthropic insiders publicly worry about their own models.
His proposed fix instead of voluntary restraint: cross-testing. Anthropic runs its safety test harness on OpenAI models, OpenAI tests Anthropic, SpaceXAI tests both, and leading Chinese AI companies join the same system — everyone tries to break everyone else's models before release. The full discussion is in the latest All-In episode, alongside Starship, space data centers, and Terafab topics.
Related event: Musk Proposes AI Labs Cross-Audit Each Other's Safety Tests(2 posts)→
More from AGI Musings
- Following the money behind the "slow down AI" movement: a documented influence network — examachine · 2026-09-15
- AI risk take: robots and viruses are manageable — culture shock is the real threat — mayfer · 2026-09-15
- The AI information gap with the real world gets weirder every week — prasenx · 2026-09-15
- One static weight set for all inference? Why a many-model world looks like human society — willcb · 2026-09-15
- Capability went vertical, instrumental drives never showed up: 2008 AI-drives prediction still unfulfilled — banteg · 2026-09-15
- Reddit essay pushes back on AI doom: true superintelligence wouldn't take irreversible gambles — ProxyLumina · 2026-09-15