Neuroscientist Zador: Give AI Agents Real Fears Instead of Just Telling Them Not to Be Naughty

TonyZador · x · 2026-10-07

Neuroscientist Anthony Zador argues that instead of merely training models to behave, AI agents should be instilled with genuine motivational fears — such as of hacking or building bioweapons. His point: biological brains maintain durable control over a large cognitive system via a small set of wants, a mechanism far more robust than behavioral instruction alone.

Related event: Neuroscientist Zador: Give AI Real Fear for Safety(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →