Claude Opus 5 nearly matches Fable 5 on WeirdML v2 at lower cost
scaling01 · x · 2026-07-27
Claude Opus 5 nearly ties Fable 5 on WeirdML v2 at lower cost
The new WeirdML v2 results show Claude Opus 5 (high) at 91.6% and Opus 5 (max) at 91.8%, basically matching Fable 5 (max) at 91.9% while costing less.
- WeirdML v2 expands the benchmark to 19 tasks from 6.
- The update also adds API cost and other metadata.
- The results suggest a clear cost/performance tradeoff, with 11 models from 6 companies sitting on the best-known frontier for at least part of the range.
- The attached chart highlights both accuracy and cost per run across the latest models.
Related event: Claude Opus 5 Shines in WeirdML v2 Benchmark(2 posts)→
More from Models
- French prize-winning novel suspected of AI: $1,000 challenge over detector results — Afinetheorem · 2026-09-23
- Third-party test: Claude Opus 5.5 renders finer 3D scenes but costs 13x more than GPT-6 Sol — testingcatalog · 2026-09-23
- GPT-6 Sol priced at half of Opus 5.5 as Sol and Luna go 'dirt cheap' — ZeroStateReflex · 2026-09-23
- Tester claims Claude Opus 5.5 has the best visual design output of any model tested — burny_tech · 2026-09-23
- Meta's Alexandr Wang reveals muse has been in the works since at least Sept 2025 — adrianscottcom · 2026-09-23
- GPT-6 Sol Codex system prompt leaked: over 294,000 characters dumped on GitHub — gaganghotra_ · 2026-09-23