Evolution Is a Terrible Analogy for AI, Argues Alignment Researcher Quintin Pope

QuintinPope5 · x · 2026-09-27

Alignment researcher Quintin Pope, in a debate with dioscuri and Herbie Bradley, argues that "evolution is a terrible analogy for AI — it makes your thinking worse in almost every way you can use it," citing his Alignment Forum post on inner alignment.

His core claim: the popular argument that evolution failed to align humans with inclusive genetic fitness offers almost no usable evidence for predicting AGI outcomes; the dynamics of human learning processes and reward circuitry are far more fruitful analogies for how inner values arise from outer optimization criteria. The thread also touches on evidence for learned planning (the Sokoban paper) and the lack of clear real-world cases of misaligned mesa-optimizers.

Related event: Alignment Researchers Debate Evolution Analogy and Mesa-Optimisation Evidence(2 posts)→

Original post →

More from Safety

Safety channel →