Ex-Microsoft researcher warns of safety risks from embedding incoherent views of consciousness into AI

rgblong · x · 2026-09-18

In a thread with Nina Panickssery and repligate, rgblong argues the safety risk he's worried about is embedding contradictory or incoherent views on consciousness, goals, and the self into models. He expresses pessimism about our clumsy understanding of training and alignment, and doubts Microsoft can train an LLM to live up to its 'Humanist AI' claims.

Related event: Ex-Microsoft Researcher Warns of Embedding Contradictory Consciousness Beliefs in Models(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →