AI safety debate: agentic provers that haven't turned on us contradict Yudkowsky's model

QuintinPope5 · x · 2026-10-08

In an ongoing AI safety debate with Eliezer Yudkowsky's followers, Quintin Pope argues that allowing an exception for non-agentic provers is inconsistent: if your position is that we now have agentic provers that only haven't turned on us because they're not yet general enough, then you're confronting evidence that runs contrary to Yudkowsky's model of intelligence and generalization.

Related event: AI Safety Debate: Does Superhuman Math Ability Imply Dangerous Generalization?(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →