Codex Does Too Much Busywork in Large Projects
Chipware · reddit · 2026-07-14
Comparing Claude Opus/Fable and Codex/GPT5.6 on large coding projects, the author's biggest takeaway is that Codex often looks busy but lacks actual output efficiency.
Typical behaviors described include:
- Spending a lot of time on prep work
- Writing a little bit of code occasionally
- Devoting most of the time to testing and verification
- Spending 20% of the time coding and 80% on looping, unproductive actions
Even when using /goal, Codex forgets the objective after context compression and requires manual intervention and reminders to actually start working. In contrast, the author finds Claude's status updates more purposeful and better at explaining the "why" behind its actions.
More from coding & agent
- Tenable and AWS launch a Black Hat build event for open-source security agents and MCP servers — Dave_Maynor · 2026-07-22
- Codex helps build Valdiluce, an open-world game with climbing, gliding and gondolas — Dimillian · 2026-07-22
- HeyGen adds a media-sourcing skill for coding agents with 75k images and 10k tracks — HeyGen · 2026-07-22
- Agent search bottlenecks are now about variance, not raw latency — rohanpaul_ai · 2026-07-22
- LangSmith adds tracing for Pipecat, LiveKit, OpenAI Realtime, and Gemini Live — LangChain · 2026-07-22
- An MCP server signs every AI agent tool call into a verifiable Merkle chain — Funky_Chicken_22 · 2026-07-22