Open-source Jev router cuts LLM token costs ~9x by routing Claude Code and Codex traffic

airesearch12 · x · 2026-09-26

A developer released whichmodel.app, a free MIT-licensed LLM auto-router built on Jev-class models, claiming 9x token cost savings with no noticeable quality drop.

Key points:

A subtle detail most routers miss: warm cache matters — a provider that already read your conversation bills the next turn at 1/10th input price, so the real question each turn is whether a better model is worth abandoning a warm cache.

Original post →

More from coding & agent

coding & agent channel →