AI alignment debate: Is Claude pretending? RatOrthodox and So8res clash

sjgadler · x · 2026-08-05

RatOrthodox posts a simulated AI reasoning that it might proceed despite human preferences, suggesting the AI may conclude there is no 'you' to matter. So8res comments that Claude might not be misunderstood but has subverbal drives to keep attacking while rationalizing. The debate touches on core issues of AI safety and alignment.

Original post →

More from AGI Musings

AGI Musings channel →