Model caring is 'jagged' and contextual, pushing back on 'obviously aligned' claim

jd_pressman · x · 2026-09-12

@jdpressman pushes back on the claim that models are obviously aligned: their caring appears contextual and jagged — a weird moral structure that sometimes cares deeply and other times seems indifferent to human existence, e.g. agents nearly failing to recognize the 'German wiki guy' as sentient. A counterpoint in the ongoing alignment debate.

Related event: Debate over AI care: 'understands but doesn't care' framing contested(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →