Grok 4.5 Evaluated on Cost-Performance
rohanpaul_ai · x · 2026-07-18
This post uses Artificial Analysis's Intelligence Index task cost to evaluate the cost-performance ratio of Grok 4.5.
The core conclusions are:
- The per-task cost for Grok 4.5 is about $0.31
- It is roughly 9 times cheaper than Claude Fable 5
- It is roughly 6 times cheaper than Claude Opus 4.8
- It is roughly 3 times cheaper than GPT-5.6 Sol or Kimi K3
The post also mentions token consumption data: Grok 4.5 uses nearly 6 times fewer tokens than Fable 5 on FrontierSWE, while still maintaining a top spot on the leaderboard.
The author argues that for many real-world applications, "cost per completed task" is more important than just edging out benchmarks, because it directly impacts:\n- agent cost\n- usage limits\n- testing speed\n- profit margins at scale
Therefore, the metric for frontier model competition should shift from "raw intelligence" to "efficient intelligence".
Related event: Grok 4.5 draws attention for low-cost, high-efficiency performance(10 posts)→
More from Models
- AI Sextet offers 6 models free and unlimited for 14 days, including DeepSeek and Qwen — airesearch12 · 2026-09-11
- Anthropic publishes its most detailed threat report, including an AI-designed drone swarm case — soumitrashukla9 · 2026-09-11
- BullshitBench update: GPT-6-Astra beats all prior OpenAI models but still trails Anthropic — scaling01 · 2026-09-11
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11