OpenAI boosts default GPT-6 Astra and GPT-6.1 Sol speed ~50% to 50 tokens/sec
kimmonismus · x · 2026-10-06
OpenAI has optimized default inference speed by roughly 50% across GPT-6 Astra and GPT-6.1 Sol, bringing them to about 50 tokens per second. The improvement applies through the subscription layer to every product and partner using Sign in With ChatGPT — including OpenCode, Pi, Amp, and Devin — with no changes needed on the user side and rollout within two hours. Commentator kimmonismus argues this speed bump is preferable to a model reset.
Related event: OpenAI Boosts GPT-6 Default Speeds by ~50%(2 posts)→
More from Models
- Questioning the launch: new model mirrors PrismML scales and kernels without attribution — _xjdr · 2026-10-06
- $3,500 Blackwell Personal AI PC: RTX PRO 4000 Runs Qwen Next at 50-70 tok/s — Jackyhuang · 2026-10-06
- Reddit user: MiniMax H3 and ref mods are "really incredible" — CompleteBed1797 · 2026-10-06
- Reddit users report sudden wave of refusals from Claude with no clear cause — astrorocks · 2026-10-06
- Bought an RTX 5060 for local LLMs — complex tasks scored 2/10 vs 9/10 in the cloud — Tricky-Brother-7 · 2026-10-06
- Full-bandwidth Transformer paper revised: latent feedback nears 1.5x-token performance — _arohan_ · 2026-10-06