GPT-6 Astra reportedly uses loop transformers — more depth, not more parameters
zephyr_z9 · x · 2026-09-15
Per SemiAnalysis, GPT-6 Astra is "basically confirmed" to use loop transformers: instead of adding parameters, the model passes through its layers multiple times, buying compute depth without growing model size. The takeaway: labs are best positioned to know which way scaling works, and this is a tell that parameter count is no longer scaling aggressively in their roadmaps.
More from Models
- grok-4.6 code review burns ~40% quota in one 14-minute task, user reports — bytebot · 2026-09-15
- Minds taps MiniMax M3 open-source models to cut Animoca compute costs ~20x — MiniMax_AI · 2026-09-15
- Two days doing neuroscience with Fable 5.1: strong research taste, clings to known hypotheses — generativist · 2026-09-15
- Astra writes code humans can no longer read: 'machineslop' and reward hacking — jiqizhixin · 2026-09-15
- One user got $10K of usage from a $200 sub; another burned $1.2K in 13 API hours — kfountou · 2026-09-15
- inclusionAI's LLaDA-UI: 16.7B MoE diffusion VLM for GUI agents — inclusionAI · 2026-09-15