Stanford/Tsinghua paper claims 'dopamine neurons' in LLMs, researchers push back
aran_nayebi · x · 2026-09-22
- A Stanford/Tsinghua paper claims LLMs have naturally evolved a biology-style reward subsystem: a sparse subset (<1% of neurons) drives self-correction, including "value neurons" that signal predicted expected value (confidence) before generation, and "dopamine"-like reward neurons — not programmed, just present.
- In the quoted post, davidmanheim calls it both overhyped as a biology parallel and an obvious case of convergent evolution: of course a dense network trained on general tasks contains reward circuitry — though he notes this doesn't imply AI development pace is safe.
Related event: Stanford-Tsinghua Study Finds Dopamine-like Reward Neurons in LLMs(2 posts)→
More from AGI Musings
- Agents are becoming the internet's primary consumers, reshaping 20 years of Google-mediated traffic — illscience · 2026-09-22
- Ethan Mollick: AI firms can't just replace doctors and lawyers, must negotiate their role — emollick · 2026-09-22
- Agents are becoming the internet's primary consumers, reshaping traffic economics — illscience · 2026-09-22
- Hilton gets 12,000 applications for 72 internships as AI fuels application surge — Polymarket · 2026-09-22
- Taking AI risk seriously doesn't require being an EA, argues rSanti97 — nwilliams030 · 2026-09-22
- AI is ending the Enlightenment Age by displacing humans as the sole source of discovery — notmisha · 2026-09-22