Grok 4.6 Nearly Matches Claude Fable 5 on Agentic Benchmark at a Fraction of the Cost
ArtificialAnlys · x · 2026-08-13
On Artificial Analysis' AA-Briefcase agentic knowledge work benchmark, xAI's Grok 4.6 made significant gains, scoring neck and neck with Anthropic's Claude Fable 5.
The benchmark evaluates models on long-horizon agentic tasks. Notably, Grok 4.6 delivers a cost per task of just $4.42, substantially cheaper than Claude Fable 5 ($22.30) and Claude Opus 5 ($17.79), demonstrating an impressive intelligence-per-dollar ratio.
Related event: Grok 4.6 Tops Agentic Benchmarks with High Performance and Cost-Efficiency(5 posts)→
More from Models
- Anthropic Slammed by Users for Extra 'Fast Mode' Fees on Claude Code — Neel_MynO · 2026-08-13
- Claude Opus 5 Responses Are 3x Longer, Simpler Vocabulary, and Loves Em Dashes — arena · 2026-08-13
- Researcher Mocks Token Overspending: 'It's Mostly a Skill Issue' Now — deliprao · 2026-08-13
- User Test: Claude 3.5 Sonnet V4 Destroys Opus in Coding at Near-Zero Cost — EAccelerate_42 · 2026-08-13
- Hands-on: Grok 4.6 is Fast, Smart, and Collaborates Across Agents — doodlestein · 2026-08-13
- xAI Launches Grok 4.6 Model and Early Beta Grok Bot Teammates — aman_madaan · 2026-08-13