Critique of alignment: solve engineering problems before proceeding

davidmanheim · x · 2026-08-31

Discussing AI alignment, the author argues we should solve underlying engineering problems before advancing, rather than repeating failed strategies. He cites discussions on why iterative alignment might fail, noting how incentivized cheating on impossible tasks and moderate misalignment can accumulate across model generations.

Related event: Failed Iterations Don't Guarantee Success in AI Alignment(2 posts)→

Original post →

More from Safety

Safety channel →