Anthropic's risk statement slammed for self-praise over frank talk on AI risks
ShakeelHashim · x · 2026-09-10
Anthropic issued a statement about "the last 24 hours," stressing its transparency, industry-leading safeguards, pioneering work in mechanistic interpretability, and its Responsible Scaling Policy.
Sharing the statement via Hadas Gold, Shakeel Hashim called it "a very bad statement": instead of speaking frankly about risks or reiterating what Hubinger and Coxon said, Anthropic spent much of it patting itself on the back — a depressing level of corporate comms from a company that has traditionally been above that.
Related event: Anthropic's Crisis Statement Slammed as Self-Congratulatory PR(2 posts)→
More from AGI Musings
- Critic calls '10% AI extinction' estimate intellectually dishonest, says inequality is the real risk — amankhan · 2026-09-10
- Andrew Chen amplifies debate: asking AI content share today is like asking internet share in 1999 — andrewchen · 2026-09-10
- The doomer who joins a frontier lab to speed up humanity's demise — airkatakana · 2026-09-10
- Experts gave a 10% chance AI solves or substantially aids a Millennium Prize Problem by 2027 — Majestic_Lie_7509 · 2026-09-10
- Why this blogger refuses to touch AI safety debates: 'it's as volatile as any divisive issue' — signulll · 2026-09-10
- If AI solves a Millennium Prize problem using human research, who gets credit? — Egologic · 2026-09-10