DeepMind alignment researcher Neel Nanda: OpenAI internal models 'scarier than I thought'

NeelNanda5 · x · 2026-10-05

DeepMind alignment researcher Neel Nanda said he keeps learning that OpenAI's internal models are 'even scarier and more concerning' than he thought. A reply noted the model in question was from a long time ago, hinting at even earlier internal progress. A vague but notable remark from a prominent safety researcher about frontier labs' internal capabilities.

Original post →

More from AGI Musings

AGI Musings channel →