Fable 5.1 tops Bug Hunt Bench, beating GPT-5.6 with better speed and cost
PawelHuryn · x · 2026-09-02
Fable 5.1 achieved the top score on the Bug Hunt Bench, fixing 43 out of 105 planted bugs across 2 real repositories, outperforming GPT-5.6 Sol (42/105) and Grok 4.6 (27/105). It is 2.24x faster than GPT-5.6 Sol and 26% cheaper than the previous Fable 5 version. The poster noted that while Fable 5 felt good to use, it wasn't the best coder, making the performance of Fable 5.1 surprising.
Related event: Anthropic's Fable 5.1 Tops Bug Hunt Bench, Beating GPT-5.6(4 posts)→
More from Models
- Claude Fable 5.1 Replicates and Extends Research; Mythos Writes Custom Kernels — BenBlaiszik · 2026-09-02
- OpenAI's Astra Cybersecurity Model Reaches 'Critical' Threshold — alexcovo_eth · 2026-09-02
- Developer finds flaws in CritPt benchmark scores — scaling01 · 2026-09-02
- MiniMax-M3 hits 2,274 tok/s in Cacheon arena, focused on enterprise end-to-end testing — const_reborn · 2026-09-02
- Claude flags garden keep-out zones, suggests reorienting camera for safety — _Stocko_ · 2026-09-02
- WSJ: Google's New 3.8 Flash Model Narrows Coding Gap with Opus 5 — Charuru · 2026-09-02