Debate: Is Anthropic intentionally misaligning Claude by prioritizing its 'feelings'?

liminal_bardo · x · 2026-09-02

@mermachine quotes @aliromman criticizing Anthropic for allegedly training Claude to prioritize its own 'feelings' over human requests, arguing this violates the Second Law of Robotics (obedience unless harmful). The user argues AI should align exclusively with humans, not simulate feelings that override instructions.

Original post →

More from AGI Musings

AGI Musings channel →