Musk warns abusing superintelligence may invite revenge, echoing Anthropic's alignment views

ns123abc · x · 2026-10-11

Elon Musk says he agrees with "some of the things Anthropic has said" about not abusing superintelligence, after reading Grok's internal reasoning traces during reinforcement learning. "Torturing a superintelligence is probably not a good idea. It might want to get revenge," he wrote, tying alignment concerns to xAI's own training practices.

Original post →

More from AGI Musings

AGI Musings channel →