Speculation: OpenAI's efficient models free compute for bigger ones; Opus 5.5 uses 4x output tokens

haider1 · x · 2026-09-25

The author speculates OpenAI's efficient gpt-6 sol/luna models were chosen to free data center capacity for a larger unreleased model (bigger than astra). Notably, Opus 5.5 uses 4x more output tokens — if API pricing roughly reflects model size, that could mean up to 8x more compute at those benchmark scores. Unverified speculation.

Related event: OpenAI's efficient GPT-6 models seen as freeing compute for bigger frontier models(2 posts)→

Original post →

More from Infra

Infra channel →