rasbt debunks Astra looped-transformer hype: fewer tokens likely mean smarter model, not hidden reasoning

rasbt · x · 2026-09-04

Responding to The Information's report that OpenAI's Astra uses a 'recurrent depth or looped transformer' architecture, Sebastian Raschka argues the looped design isn't hiding reasoning tokens. GPT 6 Astra using fewer tokens than GPT 5.6 Sol likely reflects it being a smarter, bigger model. He notes the same pattern within prior generations: GPT 5.6 Sol uses 46% fewer output tokens than GPT 5.6 Luna on the intelligence index, yet nobody claims Sol hides its reasoning more.

Related event: Rumors of Looped-Depth Architecture in OpenAI's Astra Spark Debate(3 posts)→

Original post →

More from Models

Models channel →