Sebastian Raschka breaks down GPT-6 Astra rumors, looped transformers and hidden chains of thought
rhiever · reddit · 2026-09-10
Sebastian Raschka published a new essay covering three topics: the rumored GPT-6 Astra, the case for looped transformer architectures, and how hidden chains of thought work in reasoning models. He offers his own technical judgment on the rumors and architecture trade-offs in his usual practitioner style.
Related event: Raschka Breaks Down GPT-6 Astra and Looped Transformers(4 posts)→
More from Models
- DeepSeek v4.1 flash preview spotted running at ~300 tok/s, vision still missing — kevinkern · 2026-09-10
- Viral rant slams OpenAI's Navier-Stokes claim: 10k agents for 88 hours forced a singularity, solved nothing — thedealdirector · 2026-09-10
- Testing computer use grounding: asking an AI to draw a portrait inside Google Calendar — xwang_lk · 2026-09-10
- OpenAI-compatible endpoint adds prepaid spend caps and per-key daily limits — testingcatalog · 2026-09-10
- User dumps Astra for Sol, burns through weekly token limit in 5 hours — StewartalsopIII · 2026-09-10
- Claude user burns through 98% of $200 Pro weekly limit with just $46 of API-equivalent usage — Angaisb_ · 2026-09-10