Debate: Do parameter counts determine frontier model pricing?
On August 30, analyst @scaling01 offered his estimates on the parameter scale and pricing of OpenAI's frontier models: he guesses that GPT-5.5/5.6, currently priced at $20 for output, has roughly 2-3 trillion parameters, and if OpenAI later releases a full 10-trillion-parameter Astra rather than a distilled version, scaling proportionally suggests its output price could reach $66-100 per million tokens.
Confirmed
- This is @scaling01's speculation based on a parameter-to-pricing ratio, not an official disclosure; the parameter counts of GPT-5.5/5.6 themselves have not been confirmed by OpenAI
- Comparison data cited in the discussion: Nemotron (550B) at $2.40/M, Trinity (398B) at $0.80/M, Qwen MoE (397B) at $3.50/M
- @xeophon added that GLM-5.3-Flash uses an architecture different from the flagship version, and that a 2-3T-scale model priced at just $20 makes GLM strong value for money
Not Yet Confirmed
- There is no official information on whether Astra exists, whether it would be a full 10T-parameter version, or its final pricing; the $66-100 figure is merely a proportional extrapolation
- @scaling01 also cautioned that comparing prices of models from the same company in the same month is insufficient; cross-model price comparisons need to account for architectural differences
Why It Matters
- The discussion links parameter scale to per-token pricing, offering a quantifiable framework for tracking how frontier models are commercially priced; if a 10T-class model actually ships, its output costs would significantly raise the barrier to high-end inference services, making the relative advantage of cost-effective models like GLM even more pronounced.
2026-08-30 ~ 2026-08-31 · 9 related posts
Primary sources
- [source] Estimate: a full 10T-parameter OpenAI Astra could cost $66–$100 per million output tokens — scaling01 · 2026-08-30
- Speculation: GPT-5.5 might be 2-3T params; a full 10T Astra could cost up to $100 — scaling01 · 2026-08-30
- Model cost analysis: 10T Astra could cost up to $100 per million tokens — xeophon · 2026-08-30
- Comparison of MoE Model Pricing: Nemotron, Trinity, Qwen, and GLM — scaling01 · 2026-08-30
- [source] Scaling01 on LLM Pricing: Linear Cost Scaling, Margin Drives Price — scaling01 · 2026-08-31
- [source] Model pricing cannot be inferred from total parameter count — xeophon · 2026-08-31
- Model params don't dictate price; OpenAI's frontier premium is about margin — xeophon · 2026-08-31
- Zhipu clarifies GLM architecture and pricing misconceptions — scaling01 · 2026-08-31
- GLM-5.3-Flash uses entirely different architecture from flagship, says insider — xeophon · 2026-08-31