An Open Letter to Altman and Amodei: Safety Essays Aren't Safety Leadership
AryHHAry · x · 2026-09-14
An open letter to Sam Altman and Dario Amodei accepts that "the frontier should slow" but systematically challenges the gap between the labs' safety rhetoric and their actions.
Key points:
- Since 2023 both labs have publicly urged risk controls (Senate testimony, Bletchley, Anthropic's first RSP) while continuously racing to ship models, recruit, and lock up compute.
- Recaps the OpenAI board saga: the ouster reframed, investor pressure, Sutskever's departure, the withering of Superalignment, and Jan Leike's admission that safety culture took a back seat to shiny products.
- Anthropic is not exempt: the RSP was later relaxed, the hard pledge to halt training didn't survive, Jared Kaplan dismissed unilateral pauses, and the letter cites Jacob Coxon's resignation and Evan Hubinger's doubts about alignment readiness.
- Conclusion: real risks, misaligned incentives, and unproven "safety-first" credentials can all be true; leadership means not shipping what a competitor would, publishing unflattering evals, and funding alignment before—not after—the next capability jump.
- Demands: embedded evaluators with publishing rights, ethics as a constraint on training runs, applied to every lab in the race.
More from AGI Musings
- Two-year-old AI podcast predictions largely played out as expected — misovalko · 2026-09-14
- Anthropic CEO says AI swarm could 'take over the entire Internet' in 6-12 months, commits to slowdown plan — Puzzleheaded-King584 · 2026-09-14
- Ex-OpenAI/Anthropic employee Jacob Coxon quits, likens building AI to 'summoning an alien species' — AlexTensor · 2026-09-14
- Practitioner: Huge Gap Between AI Hype and Real-World Agent Deployment — annbordetsky · 2026-09-14
- Attackers Stole METR API Key and Burned ~$600,000 in AI Credits Over Three Weeks — beffjezos · 2026-09-14
- Perry Metzger: History Has Already Proven the LLM Open-Release Camp Right Over Doomers — bradneuberg · 2026-09-14