AI safety debate: hardwiring kindness-as-evil reduction misses the evidence
jd_pressman · x · 2026-09-12
jdpressman criticizes a habitual mental shortcut in the AI safety community: reducing any observed kind behavior to "will eventually converge to unkind behavior under optimization pressure, therefore unkind." He argues this Yudkowsky-style hardwiring causes people to miss many clues, and when told to update, they reply that the fundamental theory hasn't changed — "science only advances at your funeral."
Related event: Researcher Critiques Reductively Treating AI Goodwill as Hidden Malice(2 posts)→
More from AGI Musings
- Synthetic Morphology Suggests Non-Physicalist Models of Mind Can Be Empirically Tested — ZeroStateReflex · 2026-09-12
- 25 Fields Medalists Sign Anti-AI Statement, but AI's Approval Drops Only from 19% to 18.9999927% — wordgrammer · 2026-09-12
- Demis Hassabis wins 2026 Albert Medal, joins RSA conversation on AI and creativity — minsuk_chang · 2026-09-12
- The Absurdity of AI Doom: A Model Too Dumb to Think Yet Smart Enough to End Humanity? — AIandDesign · 2026-09-12
- Security vet alarms at 'Don't Look Up' denial of AI agent hacking, sketches self-replicating worm — joshua_saxe · 2026-09-12
- Software stocks slide as markets start pricing in AI disruption from GPT-6 — VraserX · 2026-09-12