Testing 105 Bugs: Luna Beats Fable in Cost-Efficiency, Expensive Models Avoid Coding

PawelHuryn · x · 2026-08-01

Developer PawelHuryn shared an AI model routing strategy based on 105 hidden bugs, 9 frontier models, and 14 test runs:

Cost-efficiency: Luna at max effort fixed 33 bugs for $1.80, while Fable 5 fixed 29 for $104.49. The author notes that the most expensive model he uses never touches code.

Related event: Dev Open-Sources Bug-Hunt-Bench to Test LLM Debugging(2 posts)→

Original post →

More from coding & agent

coding & agent channel →