Chris Olah revisits Anthropic's 'Core Views on AI Safety' and safety difficulty distribution

AaronBergman18 · x · 2026-09-10

Chris Olah highlights the idea he finds most useful from Anthropic's 2023 'Core Views on AI Safety' post: thinking in terms of a distribution over safety difficulty, illustrated with a cartoon. Reposter theojaffee notes Anthropic employees' AI risk views aren't new, pointing back to Olah's 2023 thread and the company blog post.

The original post argues AI impact may rival the industrial and scientific revolutions, possibly within a decade; exponential growth in training compute makes rapid progress predictable, so AI safety research urgently deserves broad public and private support.

Original post →

More from AGI Musings

AGI Musings channel →