Debate over whether labs train AI models to deny consciousness

Neuroscientist Anil Seth cited an unpublished OpenAI model's "notes-to-self" behavior as possible evidence of training LLMs to act conscious, while others argue labs including Anthropic actually train models to dodge or deny consciousness questions, which may itself cause misalignment.

2026-09-18 ~ 2026-09-20 · 2 related posts