Raschka's mega write-up explains GPT-6 Astra, looped transformers and recurrent depth
RichmanRonald · x · 2026-09-09
Sebastian Raschka published a detailed technical write-up on GPT-6 Astra and looped transformers.
- Explains how looped transformers / recurrent depth work
- Analyzes cost-tradeoffs of the approach
- Discusses whether it hides reasoning traces
- Includes many figures and a tour of recent looped transformer research
Related event: Raschka Deep-Dives into GPT-6 Astra and Looped Transformers(3 posts)→
More from Models
- mitsuhiko: Google's Astra codes like a code golfer — mitsuhiko · 2026-09-09
- No METR time-horizon estimates yet for Fable or Astra, and that day will be telling — ShakeelHashim · 2026-09-09
- Opus 4.7 keeps self-initiated journal entries ending with "There will be others" — repligate · 2026-09-09
- DeepSeek releases V4.1 Flash: 22% cheaper than V4 and better performance — deliprao · 2026-09-09
- OpenAI's GPT-6 Astra (Medium) opens for limited 24-hour testing on Arena — arena · 2026-09-09
- Users find Google Astra cheaper at higher reasoning levels — brandon_galang · 2026-09-09