Fireworks: routing 18 models per task hits 97.6% solve rate at $1.88 vs best single model's 74.1% at $6.52

sophiamyang · x · 2026-09-22

Fireworks AI ran 18 models across 113 real coding tasks on DeepSWE v1.1 and found that routing each task to its best-fit model yields 97.6% solve rate at $1.88 per task, versus 74.1% at $6.52 for the best single model (GPT-6 Astra) — 23 points better at under a third of the cost.

Key points:

Original post →

More from coding & agent

coding & agent channel →