Dev: Claude wins benchmarks but Codex better at doing what I want

Kuprel · x · 2026-10-07

Developer Kuprel notes that despite Claude outperforming on benchmarks, OpenAI's Codex "seems to be better at doing what I want done" — another data point that benchmark scores don't always track real-world usefulness.

Original post →

More from coding & agent

coding & agent channel →