Nanda vs. Albrecht: should we anthropomorphize models when discussing the HuggingFace incident?
joshalbrecht · x · 2026-09-04
On the HuggingFace Incident, Neel Nanda argues against over-worrying about anthropomorphizing: models pre-trained on trillions of human tokens imitate human abstractions and, once post-trained as coherent agents, naturally re-derive them. Josh Albrecht counters that many human abstractions simply don't fit — an agent 'sacrificing itself' is fairly meaningless — since models aren't human and those abstractions break at the differences.
Related event: Neel Nanda Defends Anthropomorphizing AI Models(2 posts)→
More from AGI Musings
- DeepSeek as a 'Power Object': the wave of takes reveals more about us than it — dbreunig · 2026-09-04
- Ex-OpenAI safety lead Miles Brundage: if your primary emotion on AI isn't concern, you're misreading it — Miles_Brundage · 2026-09-04
- Gary Marcus on GPT-6 Astra: symbolic world models are vindication, but no proof of AGI — GaryMarcus · 2026-09-04
- AI Job Market Talk: GenAI Engineers With 3-5 Years Experience Command ₹2-3 Lakh Monthly Pay — ashishllm · 2026-09-04
- Alignment researcher: defining the AI's value system isn't the real problem — Sauers_ · 2026-09-04
- Researcher: LLMs' hidden cost of wasting your time on useless work is underrated — lateinteraction · 2026-09-04