AI Safety Criticism: Accusations of Performative Concern

davidmanheim · x · 2026-08-30

Amidst a debate about solving metric optimization in AI alignment, David Manheim fiercely criticizes "pro-safety" individuals who get upset when others attempt to solve the problems they warn about, suggesting they prefer lecturing over solutions. He cites his paper on variants of Goodhart's Law to explain why metric hacking approaches fail.

Related event: AI Safety Researchers Debate Goodhart's Law in LLM Alignment(4 posts)→

Original post →

More from Companies & People

Companies & People channel →