rasbt debunks Astra looped-transformer hype: fewer tokens likely mean smarter model, not hidden reasoning
rasbt · x · 2026-09-04
Responding to The Information's report that OpenAI's Astra uses a 'recurrent depth or looped transformer' architecture, Sebastian Raschka argues the looped design isn't hiding reasoning tokens. GPT 6 Astra using fewer tokens than GPT 5.6 Sol likely reflects it being a smarter, bigger model. He notes the same pattern within prior generations: GPT 5.6 Sol uses 46% fewer output tokens than GPT 5.6 Luna on the intelligence index, yet nobody claims Sol hides its reasoning more.
Related event: Rumors of Looped-Depth Architecture in OpenAI's Astra Spark Debate(3 posts)→
More from Models
- Grok web gets a cleaner redesign with a refreshed UI — XFreeze · 2026-09-05
- eyebench author says no v4, moving on to harder benchmarks — adonis_singh · 2026-09-05
- Astra-max claims vastly better intelligence-per-token even at low reasoning — adonis_singh · 2026-09-05
- Astra-max hits 95% on eyebench-v3 at half the cost of Sol-max, tokens ~3.8x fewer — adonis_singh · 2026-09-05
- pass@1 dead even, but Fable 5.1 wins pass@k over GPT 6 Astra — zainhas · 2026-09-05
- GPT-6 Astra's cybersecurity classifier blocks a non-security bug-hunting task in Codex — Elctsuptb · 2026-09-05