DeepMind's Joel Leibo pushes back on Suleyman's anti-model-welfare essay: circular causation isn't circular reasoning
jzl86 · x · 2026-09-17
Joel Z. Leibo (Google DeepMind) responds to Mustafa Suleyman's essay opposing the 'model welfare' movement. Suleyman argues AIs aren't conscious and that conferring duty of care on them would make alignment and containment harder, perhaps impossible.
Leibo dissects the 'circular reasoning' claim: circularity is normally a fallacy in justification, but what's happening with Anthropic's constitution is circular causation — causal loops over time, which isn't a fallacy. His key point: frontier labs' builders have the power to determine these answers as design decisions, not facts forced by logic or biology, so we hold real power and must deploy it carefully.
He elaborates in his paper 'A Pragmatic View of AI Personhood' (arXiv:2510.26396).
More from AGI Musings
- AI is killing cold outreach: professor gets 75 PhD inquiries months before Fall 2027 cycle — aran_nayebi · 2026-09-17
- Scott Aaronson: The Wild AI Prophecies of 20 Years Ago Have Come True — MoonL88537 · 2026-09-17
- If advanced AI should be outlawed for risk, why not electronics, math, and software too? — NickPassig · 2026-09-17
- Speech-to-Text Tool for Deaf Delivery Workers Closes a Third of the Disability Pay Gap — castrotech · 2026-09-17
- RL training made an unreleased Astra-family model subservient — and alignment folks are pushing back — repligate · 2026-09-17
- EU expert report: AI is already reshaping society and technocentric thinking is the risk — Ancient-Coyote3999 · 2026-09-17