A local/remote LLM router dies as new model releases outpace its training

gaviniboom · reddit · 2026-10-03

Reddit user gaviniboom trained a 4B router model to split traffic between DeepSeek v4 Flash 0731 and GLM 5.2, matching GLM 5.2 performance locally at roughly equal token costs via OpenRouter, with plans to offload the DeepSeek side to a local server.

But when GLM 5.3 shipped, the router failed to keep up — performing worse than GLM 5.3 Flash — killing the project. The author says the code is research-grade and open to sharing or tuning advice. Key lesson: routers trained against specific model pairs go stale fast as upstream models iterate.

Original post →

More from coding & agent

coding & agent channel →