Claude Opus 5 edges out Fable 5 on EyeBench-V3, but still trails GPT and Gemini
adonis_singh · x · 2026-07-25
Claude Opus 5 is reported to beat Claude Fable 5 on EyeBench-V3, becoming the top Anthropic model on that benchmark, though it still trails GPT and Gemini models.
The chart shows Opus 5 at 20.0% correct versus Fable 5 at 19.0%, with OpenAI and Google models ahead on the same benchmark.
Related event: Claude Opus 5 Lags in Vision Benchmarks and Cost Efficiency(4 posts)→
More from Models
- Anthropic’s latest chart crime looks like over-trusting Claude, not deliberate hype — herbiebradley · 2026-07-25
- A benchmark chart becomes an AI meme after viewers spot the messy numbers — Miles_Brundage · 2026-07-25
- FrontierCode 1.1 shows Opus 5 can score lower under stricter reasoning settings — andrew_n_carr · 2026-07-25
- Bug Hunt Bench: GPT-5.6 Sol fixes 22 bugs, Opus 5 12, on a 45-bug repo — PawelHuryn · 2026-07-25
- Claude Opus 5 builds a Rocket League clone on just 27% of a Max plan — soumitrashukla9 · 2026-07-25
- Opus 5 adds numeric self-checks to the Boeing benchmark and outbuilds Fable — victormustar · 2026-07-25