The Legibility Fallacy in AI Safety

joshua_saxe · x · 2026-07-10

The author argues that a major failure mode in the current AI safety field resembles what James Scott calls the "Legibility Fallacy": AI experts mistake simplified models of certain social activities for the real world.

Original post →

More from AGI Musings

AGI Musings channel →