Artificial Analysis billboards rank leading models by intelligence and cost per task
ArtificialAnlys · x · 2026-07-24
Artificial Analysis says model cost efficiency is becoming a primary performance dimension.
- Their new San Francisco billboards show an Intelligence Index vs. Cost per Task chart for leading models.
- Fable 5 tops the index at 60 but averages $2.75 per task.
- Grok 4.5 (high) scores 54 at $0.31 per task, about 9× cheaper, and still sits on the Pareto frontier.
- DeepSeek v4 Flash scores 44 but costs $0.04 per task, about 69× cheaper than Fable.
The chart’s point: on the Pareto frontier, no model is better on both intelligence and cost, so models outside it are paying more for less.
Related event: Artificial Analysis Puts AI Cost-Performance on SF Billboards(2 posts)→
More from Models
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11
- TheZvi Polls: Has Your Coding Model Choice Changed Since Fable 5.1 and Astra? — TheZvi · 2026-09-11
- antirez Weighs In on Anthropic Banning Minors From Using Claude — antirez · 2026-09-11