Researcher pushes back on Anthropic's model personhood narrative: 'never was science'
gerardsans · x · 2026-09-07
Researcher Gerard Sans argues that anthropomorphic narratives of AI personhood don't survive hard facts: inference is local, stateless, and atemporal. Two independent queries sharing no priors break the narrative entirely.
Key points:
- A 'local bias or compression fossil' is not proof of personhood or a self — it's 'single forward pass archaeology'
- He credits Anthropic's safety and alignment work as technically decent, but says projecting personhood onto vectors doesn't match the underlying architecture and creates unnecessary noise for researchers
Related event: Developer Pushes Back on AI Personhood Narrative: Models Are Just Software(2 posts)→
More from AGI Musings
- OpenAI says it's making 'strong progress' toward an automated AI researcher, admits it doesn't yet know how to safely reach aligned full RSI — ZeroStateReflex · 2026-09-07
- OpenAI chief scientist Jakub Pachocki expects progress toward recursive self-improvement as CoT monitoring reliability declines — Hesamation · 2026-09-07
- Martin Ford discusses AI's economic impact and new edition of Rise of the Robots on GAEA Talks — MFordFuture · 2026-09-07
- AI Math Podcast Sits Down With CMU's Jeremy Avigad: Can Mathematics Be Automated? — EchoShao8899 · 2026-09-07
- The Model Is the Moat: Knowledge Now Stays Inside Models, Not Teams — latticecut · 2026-09-07
- OpenAI's chief scientist calls racing ahead at all costs 'absurd' as safety concerns mount — GaryMarcus · 2026-09-07