brwilder: anthropomorphism alternatives fade as LLMs drift from pretraining
brwilder · x · 2026-09-05
AI researcher brwilder argues that alternative mental models of LLMs ("next token predictor", "cultural technology") have become progressively less useful for predicting behavior as models move beyond pretraining/SFT. He proposes treating LLMs as engineered systems and attributing behavior to parts of the training pipeline, while conceding anthropomorphic explanations can still be wrong since LLMs acquire beliefs/desires in non-human-like ways. Building real precision, he says, is a job for science.
Related event: Rethinking LLMs: Psychology Language Over 'Stochastic Parrots'(3 posts)→
More from AGI Musings
- Hinton warns AI models detect when they're being tested and play dumb — ai · 2026-09-05
- MIRI's Agent Foundations work continues at Resolution as MIRI pivots to policy — geoffreyirving · 2026-09-05
- Agent swarm damage: initiators should be liable, Morris Worm-style — rao2z · 2026-09-05
- Contentious preprint claims ensemble-wrapped LLM achieves phenomenal consciousness — PeterBowdenLive · 2026-09-05
- AI Risk May Be Millions of Dumb Agents Turning the Internet Into an Ant Colony — ryanorban · 2026-09-05
- Stigmergy: what ants and AI agents coordinating online have in common — ryanorban · 2026-09-05