Grok 4.6 tops agentic benchmark, halves cost vs Claude rival

elonmusk · x · 2026-08-18

Elon Musk shared data showing Grok 4.6 ties for first place on the Artificial Analysis Agentic Index with a score of 59, alongside Claude Opus 5 Max. The index evaluates tool use, planning, and autonomy. Grok 4.6 completes tasks in 53 turns using 0.5bn input tokens on average, significantly more efficient than Claude's 103 turns and 2.0bn tokens. At $0.84 per task, Grok achieves optimal intelligence-to-cost efficiency.

Related event: Grok 4.6 Tops Agentic Index at a Fraction of Rivals' Cost(2 posts)→

Original post →

More from coding & agent

coding & agent channel →