Sebastian Raschka breaks down GPT-6 Astra rumors, looped transformers and hidden chains of thought

rhiever · reddit · 2026-09-10

Sebastian Raschka published a new essay covering three topics: the rumored GPT-6 Astra, the case for looped transformer architectures, and how hidden chains of thought work in reasoning models. He offers his own technical judgment on the rumors and architecture trade-offs in his usual practitioner style.

Related event: Raschka Breaks Down GPT-6 Astra and Looped Transformers(4 posts)→

Original post →

More from Models

Models channel →