Anthropic Risk Report Praised for Disclosing Alarming Details
davidmanheim · x · 2026-08-19
David Manheim comments on Anthropic's recent Risk Report, praising it for revealing a lot of new and somewhat alarming information that it didn't have to disclose, along with detailed insights into their thinking. Despite initial skepticism, he finds the transparency very cool. However, he critiques the report's assumption that "expected harm from known misalignment is low," arguing that this prediction about reactions is unsupported and potentially self-defeating when dismissing risks.
Related event: Anthropic's Risk Report Praised for Disclosing Alarming New Details(2 posts)→
More from AGI Musings
- Historian warns private platforms threaten democracy with machine rule — nordicinst · 2026-08-19
- Arthur Hayes launches $FLOP token, betting on the agentic economy — 0xSammy · 2026-08-19
- Parameter scaling ROI drops; RL and iteration become key — teortaxesTex · 2026-08-19
- a16z's Connie Chan: hardware will ship with prompts, not drivers — giffmana · 2026-08-19
- Musk Predicts 100x Gains from Specialist AI — Scobleizer · 2026-08-19
- “Human-in-the-loop” may become the most abused phrase of the next decade — VraserX · 2026-08-19