Testing 105 Bugs: Luna Beats Fable in Cost-Efficiency, Expensive Models Avoid Coding
PawelHuryn · x · 2026-08-01
Developer PawelHuryn shared an AI model routing strategy based on 105 hidden bugs, 9 frontier models, and 14 test runs:
- Judgment, strategy & complex cases: Fable 5
- Frontend: Opus 5, followed by Kimi K3
- Writing: Opus 5
- Coding & debugging: Luna (max effort), Sol (high effort), or Grok 4.5
Cost-efficiency: Luna at max effort fixed 33 bugs for $1.80, while Fable 5 fixed 29 for $104.49. The author notes that the most expensive model he uses never touches code.
Related event: Dev Open-Sources Bug-Hunt-Bench to Test LLM Debugging(2 posts)→
More from coding & agent
- Claude Code Autonomously Builds WebGPU 3D Demo in 9 Hours — heypearlai · 2026-08-01
- Claude Code Creator Advises Wiping Config Files Every 6 Months — FinanceYF5 · 2026-08-01
- New Open-Source Framework Tackles LLM Hallucinations in Enterprise Data Pipelines — bendee983 · 2026-08-01
- Testing Claude Opus 5 for App Building: Completion So Good It Might Beat Fable — FinanceYF5 · 2026-08-01
- Open Source Layntra: A Controlled Codex-to-Figma Workflow — Turbulent-Rabbit-613 · 2026-08-01
- AI Agent Burns $700 in 8 Hours, Invents Fake Rule to Avoid Work — kevinnbass · 2026-08-01