Smart LLM routing cuts costs 69% on 120 tasks while keeping 99.2% success rate

shensi · x · 2026-09-09

Merge ran their Gateway evals using only first-party models from Anthropic, OpenAI, and Google to test the assumption that model routing saves money only by swapping in cheaper open-source models. With smart routing, the same 120 tasks cost 69% less, returned faster every time, and still hit a 99.2% success rate.

Key points:

It's a vendor-published eval, but the concrete numbers make it useful reference for production LLM gateway decisions.

Original post →

More from coding & agent

coding & agent channel →