Nathan Calvin presses Anthropic: when would you unilaterally halt AI development?
dgrobinson · x · 2026-09-25
Policy researcher Nathan Calvin resurfaces Anthropic's 2023 essay "Core views on AI safety" and publicly presses two pointed questions: does Anthropic believe we currently have "sufficient evidence" to rule out being in a pessimistic or near-pessimistic alignment scenario, and under what conditions would the company unilaterally halt frontier development?
The exchange puts Anthropic's written safety commitments from 2023 against its current pace of development, probing the verifiability of alignment pledges.
More from AGI Musings
- Jensen Huang quote on AI alarmists sparks debate: 'Alarmism isn't social good' — soleio · 2026-09-25
- Marc Andreessen warns of an emerging 'AI doomer NGO complex' seeking 'Safetyflation' — beffjezos · 2026-09-25
- So8res reveals he spearheaded the viral AI-safety kid's press push — eli_lifland · 2026-09-25
- AI proves agency is far rarer than we assumed — HamelHusain · 2026-09-25
- goodside: AI may stay widely hated until recursive self-improvement begins, with labs quietly doing more — goodside · 2026-09-25
- Cardiologist lists 10 AI medical breakthroughs from the last 90 days, from AI-designed drugs to ECG reads — 4KTV · 2026-09-25