Grok 4.5 Uses Significantly Fewer Tokens

XFreeze · x · 2026-07-09

According to a post, Grok 4.5 consumes far fewer output tokens on SWE Bench Pro compared to Claude Opus 4.8 max. It averages 15,954 output tokens versus 67,020 for the latter, roughly a 4.2x reduction. The author also reported a response throughput of 80 TPS, emphasizing lower inference costs, faster responses, and better scalability.

Related event: xAI Launches Grok 4.5: Coding and Agent Focus to Rival Opus(55 posts)→

Original post →

More from Infra

Infra channel →