Looped transformer is no dark art: rasbt debunks the OpenAI Astra rumor

rasbt · x · 2026-09-02

The Information reported that OpenAI's Astra uses a "recurrent depth / looped transformer" and may obscure chain-of-thought. rasbt offers a technical debunk:

His conclusion: Astra may be a good model, but the looped transformer aspect is a tiny architectural tweak. Reusing layers does not itself suppress visible chain-of-thought; if reasoning is hidden, the plausible mechanism is fewer explicit reasoning tokens via more recurrent passes—or the journalist misunderstood.

Related event: OpenAI's Astra Reportedly Uses Recurrent Depth for Latent-Space Reasoning(81 posts)→

Original post →

More from Models

Models channel →