A week using Claude Code, Codex, and Gemini CLI showed the same repo-breaking patterns
AIcademy-academy · reddit · 2026-07-25
A week-long side-by-side test of Claude Code, Codex, and Gemini CLI on the same real repository.
- Claude Code held multi-step context best, but became expensive when sessions drifted.
- Codex was the most literal and reliable for well-specified tasks, but struggled with ambiguity.
- Gemini CLI handled the largest context windows and the free tier made it easy to experiment, but results were less consistent.
The poster’s main takeaway is that process matters more than tool choice: keep a context file, break work into small scoped steps, and ask for a plan before edits.
More from coding & agent
- Dev builds interactive 3D product experience with GPT-6 Astra + Hyper3D Rodin — nikola_mr64990 · 2026-09-11
- Codex tip: use Sol with Astra and Luna sub-agents to save usage — pvncher · 2026-09-11
- agents-best-practices: a provider-neutral Agent Skill for designing and auditing agentic harnesses — tom_doerr · 2026-09-11
- Cognition's SWE-2 uses a KKT duality argument in RL to shift the effort Pareto curve — YouJiacheng · 2026-09-11
- First-ever Three.js Conference lands in Paris, with a panel on AI-shortened design workflows — OdinLovis · 2026-09-11
- Data engineering, not agent frameworks, is the real bottleneck for enterprise AI agents — dhruv2038 · 2026-09-11