Anthropic's Fable 5.1 Tops Bug Hunt Bench, Beating GPT-5.6
Anthropic's Fable 5.1 topped the Bug Hunt Bench by fixing 43 of 105 injected bugs across two real codebases, beating GPT-5.6 Sol (42) and Grok 4.6 (27), and became the first Pareto-optimal model—faster and cheaper while leading in performance.
2026-09-02 ~ 2026-09-02 · 4 related posts
- Fable 5.1 Tops Bug Hunt Bench, Outperforming GPT-5.6 — PawelHuryn · 2026-09-02
- Fable 5.1 tops Bug Hunt Bench, beating GPT-5.6 with better speed and cost — PawelHuryn · 2026-09-02
- Anthropic's Fable 5.1 tops coding benchmark, reaching Pareto frontier — PawelHuryn · 2026-09-02
- Bug Hunt Bench details: Data and receipts for Fable 5.1 vs competitors — PawelHuryn · 2026-09-02