Blind test: Sonnet 5.5 matches Opus 5.5 on a from-scratch Rust decompressor at 1/4 the price

_Duex · reddit · 2026-09-29

The author benchmarked Sonnet 5.5 vs Opus 5.5 (plus GPT-6 Sol/Luna) on the same self-contained task: write a dependency-free, unsafe-free DEFLATE/zlib decompressor in Rust, graded blind and fully automated against zlib with 4055 hidden tests (real streams, hand-built edge cases, 4000 corrupted inputs, a speed test, a zip bomb).

Opus earned its price on polish—zlib-style two-level Huffman tables, buffer discipline, and a self-written differential tester. For well-defined tasks, Sonnet 5.5 at medium looks like the better deal; Opus may still lead on open-ended debugging. Single task, single run per model.

Original post →

More from coding & agent

coding & agent channel →