Dev benchmark: Astra beats sol and fable on LOC, API cost and code quality in 10-15 feature test

robleclerc · x · 2026-09-19

Rob Leclerc shares his method for evaluating new models: have different models at the same thinking level implement the same 10-15 small features, then judge the results.

Astra clearly beat both sol and fable on lines of code, implied API cost, and quality — in several cases sol wrote 10x as much code as Astra.

Quoting theo: for real-world code work, the benefits of Fable and Astra massively outweigh the cost, though the benefit isn't better code per se — it's something subtler.

Original post →

More from coding & agent

coding & agent channel →