AI Safety Criticism: Accusations of Performative Concern
davidmanheim · x · 2026-08-30
Amidst a debate about solving metric optimization in AI alignment, David Manheim fiercely criticizes "pro-safety" individuals who get upset when others attempt to solve the problems they warn about, suggesting they prefer lecturing over solutions. He cites his paper on variants of Goodhart's Law to explain why metric hacking approaches fail.
Related event: AI Safety Researchers Debate Goodhart's Law in LLM Alignment(4 posts)→
More from Companies & People
- Musk envisions everyone owning a personal C-3PO: one person directing thousands of Optimus robots — elonmusk · 2026-09-23
- Department shut down after engineers secretly pooled access to forbidden AI models via a shared folder — Sudden_Rip7717 · 2026-09-23
- Muse's success shows users yearn for walled gardens, and Meta holds all the cards — tekbog · 2026-09-23
- Investor visits Beijing and Shanghai AI labs: China's labs are far less coordinated than US framing suggests — Dan_Jeffries1 · 2026-09-23
- Mathematician Elliot Glazer debunks AI math rumors: Anthropic Millennium Problem claim 'almost certainly false' — burny_tech · 2026-09-23
- OpenAI and Anthropic Launching Models Same Day Is Peak Game Theory, Says Paras Chopra — paraschopra · 2026-09-23