Open-weight models vs Claude Sonnet 5: GLM 5.3 wins 20 real coding tasks at 1/10th the cost

shensi · x · 2026-09-15

The mergeapi team benchmarked five open-weight models — GLM 5.3, GLM 5.3 Flash, DeepSeek V4 Pro, DeepSeek V4 Flash, and Kimi K3 — against Claude Sonnet 5 on 20 real coding tasks. GLM 5.3 came out on top while costing about one-tenth of Claude's price. Full results in the linked writeup.

Original post →

More from coding & agent

coding & agent channel →