Qwen 3.8 27B Beats Codex in Coding Benchmarks: Wins 8/13, Costs 1/3

tokenbender · x · 2026-08-16

A developer tested Qwen 3.8 27B on refactoring tasks against Codex 5.4. Qwen won 8 of 13 cases, Codex 5.5, with 2 overlaps and 2 neither. Qwen caught a simple race condition Codex missed, produced better output, and cost about 1/3 as much.

Original post →

More from coding & agent

coding & agent channel →