Senior Dev Finds Claude Opus 5 Excels at One-Shots but Struggles with Legacy Codebases
markusn42 · reddit · 2026-08-14
A senior software engineer shares observations from using Claude Opus 5. They note that the model has historically been frustrating for large, existing codebases, often struggling with technical debt and conflicting documentation.
However, when tested on a brand new project (a non-trivial Chrome extension scraping utility) in high-reasoning mode, Opus 5 performed exceptionally well. It set up the repo architecture autonomously, made 17 commits in a single session, and produced working code that passed an adversarial review by Codex.
The engineer speculates that Opus 5 might be specifically bad at complex projects with legacy debt but excels at clean-slate, one-shot tasks, suggesting that current benchmarks need to cover this aspect better.
More from coding & agent
- Tencent Releases UI-Mate-27B, a Desktop GUI Agent Model — tencent · 2026-08-24
- Comparing AI Subscriptions: DeepSeek API vs. Claude Pro vs. Local LLMs — Unlikely_Bluejay5392 · 2026-08-24
- Claude Code introduces 'Remote Control' feature to boost coding efficiency — rohanpaul_ai · 2026-08-24
- rauchg lays out fx extension philosophy: MCP, Skills, Plugins and Unix composition — AccBalanced · 2026-08-24
- Netflix details its production LLM judge: hundreds of thousands of recommendations scored weekly — omarsar0 · 2026-08-24
- smolvm passes Simon Willison's Fable 5 agent test as a secure sandbox — yawnxyz · 2026-08-24