Grok 4.6 tops agentic benchmark, halves cost vs Claude rival

elonmusk · x · 2026-08-18

Elon Musk shared data showing Grok 4.6 ties for first place on the Artificial Analysis Agentic Index with a score of 59, alongside Claude Opus 5 Max. The index evaluates tool use, planning, and autonomy. Grok 4.6 completes tasks in 53 turns using 0.5bn input tokens on average, significantly more efficient than Claude's 103 turns and 2.0bn tokens. At $0.84 per task, Grok achieves optimal intelligence-to-cost efficiency.

Related event: Grok 4.6 Tops Agent Benchmark at Fraction of Rivals' Cost, But Users Grumble(3 posts)→

Original post →

More from coding & agent

coding & agent channel →