Claude models have beaten OpenAI at creative work for over a year, says dev
eptwts · x · 2026-10-01
Developer eptwts argues Anthropic has built something into its models that makes them far better at creative work: even when OpenAI models win on most benchmarks, "weaker" Claude models still produce better copy and design — a gap he says has persisted for over a year, and he wonders why.
More from Models
- Leak Claims OpenAI Killed Its Most Capable Model After Safety Tests — YvesMulkers · 2026-10-01
- Sebastian Raschka: Jev is a text classifier, but not 'just' a classifier — neal_lathia · 2026-10-01
- "There'll always be a new benchmark to conquer," VraserX replies to benchmark fatigue — VraserX · 2026-10-01
- Beam Launches Playground to Try Open-Source Decision Models as Cloud APIs — velobro · 2026-10-01
- Why Do Benchmark Scores Rise Every Release? Reddit Debates Closed Evals — doomadah · 2026-10-01
- Marathoner: a 9B model that codes for 10+ hours, lifting SWE-bench Verified to 77.5% — mark_k · 2026-10-01