GPT-6 is said to be near, with OpenAI betting on faster inference and custom chips
haider1 · x · 2026-07-22
GPT-6 is framed as not far off and likely to be both more capable and more efficient. The post says OpenAI has invested heavily in infrastructure, custom chips, and faster inference, and that upcoming Cerebras upgrades will push speeds beyond 750 tokens per second.
It also points to GPT-5.6 as evidence: the model is described as delivering “fable-level” performance while being faster and costing about half as much.
More from Infra
- Hugging Face says GLM 5.2 handled forensic analysis after hosted models blocked it — dotey · 2026-07-22
- Do prompt caches meaningfully cut costs for production AI agents? — MembershipEmergency7 · 2026-07-22
- Reddit user runs Nemotron Ultra 550B across aging MI50 and P40 GPU rigs — Old_Grapefruit8774 · 2026-07-22
- SK Hynix denies talks to buy Intel’s Ohio fab after market rumors — rwang07 · 2026-07-22
- Ineffable Labs takes delivery of a Vera Rubin NVL72 cluster — deanwball · 2026-07-22
- Xiaohongshu's OSDI Paper: All-Flash ANNS System Cuts 90% Costs — 机器之心 · 2026-07-22