Don't panic about latent-space reasoning: OpenAI/Anthropic already monitor activations

eliebakouch · x · 2026-09-03

Pushing back on fears that models will stop "thinking" in tokens, the author argues scaling model size always requires scaling depth for reasoning efficiency, and that OpenAI and Anthropic are deliberately training for reasoning efficiency per unit cost. Models still emit millions of reasoning tokens on hard problems, and even if reasoning shifts into latent space, monitoring will keep up — both labs' oversight systems already inspect model activations, which is literally the latent space.

Related event: OpenAI researchers push back on neuralese fears, saying frontier models remain monitorable(20 posts)→

Original post →

More from AGI Musings

AGI Musings channel →