Satire: Grok 4.6 Matches Claude Fable 5 on WANDR at One-Third the Cost

elonmusk · x · 2026-08-14

Elon Musk shared a satirical benchmark review filled with fictional model names. The tweet claims that on Perplexity's WANDR benchmark, Grok 4.6 matched Claude Fable 5 with a score of 0.496, but at a fraction of the cost ($7.58 vs $20.30 per task), while also outperforming GPT-5.6 Sol.

This is a humorous jab by the AI community at the constant benchmark racing, version number inflation, and pricing wars among current LLMs.

Original post →

More from Fun

Fun channel →