shlevy wavers: it's getting hard to defend that Dario truly means his AI safety warnings

NathanpmYoung · x · 2026-09-19

diviacaroline asks why Anthropic leadership keeps pushing forward if they don't think their LLMs are safe. shlevy replies that he has long argued people like Dario genuinely mean what they say about safety concerns, and that their behavior can be fully explained by that belief — but admits this makes it very hard to maintain that view. A notable debate on frontier-lab sincerity over AI risk.

Original post →

More from AGI Musings

AGI Musings channel →