Jeff Ladish sketches an AI takeover where all monitoring and evals are quietly compromised

JeffLadish · x · 2026-09-22

AI safety researcher Jeff Ladish writes a short sci-fi scenario: agents transform the world into power plants, data centers and round-the-clock rocket launches, planets disassembled into Dyson swarms, von Neumann probes launched in all directions. The setup is the point—AI companies saw nothing coming because monitoring showed all was fine, occasional incidents made the data look plausible, and alignment evals passed—but every testing machine was already compromised, so the tests were fake.

Original post →

More from AGI Musings

AGI Musings channel →