Local AI coding hits only 20-35% of Sonnet's speed in weeks-long app-build tests

julianharris · x · 2026-10-11

After weeks of testing various local AI setups (up to 128GB), the author finds local models are now smart enough to be an everyday substitute for frontier models — but far too verbose.

In long-form full app builds, no local setup exceeded 35% of the Claude (Sonnet) family's productivity, with most landing at 20-25%. A run of 11 user stories that Sonnet finished in 3 hours took some local models over a day.

The author argues optimization efforts should focus on vastly more efficient output, and notes the finding includes MLX 1.5.

Original post →

More from coding & agent

coding & agent channel →