Bug Hunt Benchmark: Opus 5.5 nears Fable 5.1 at 2/3 cost; GPT-6.1 Sol 10x cheaper

PawelHuryn · x · 2026-10-02

Paweł Huryn ran his Bug Hunt Benchmark (2 repos, 105 bugs frontier models missed in early 2026) to compute the real API value of AI subscription plans:

Model results

Subscription value

Related event: Bug Hunt Benchmark Tests Real Value of AI Subscription Plans(2 posts)→

Original post →

More from Models

Models channel →