GPT-4o to GPT-6.1 and Opus 3 to Opus 5 charted on the same independent benchmarks

FlorianGallwitz · x · 2026-10-10

A comparison charts GPT-4o, o3, GPT-5, GPT-5.5, GPT-6.1 Sol, and Claude Opus 3/4/4.5/5 across the same independent benchmarks, using data from Epoch AI and benchmark maintainers rather than self-reported vendor numbers, visualizing two years of frontier model progress.

Original post →

More from Models

Models channel →