Developer Test: Claude Opus Excels as Async Agent, GPT Leads in Instruction Following

brandon_galang · x · 2026-07-31

A developer shared their experience after switching to Claude full-time. They noted that while Claude used to be prone to hallucinations, adjusting prompting and verification styles has made it highly controllable today.

They found Claude Opus to be overly verbose and lacking real-time narration during work, making it harder to steer mid-task. However, it excels as an async agent for well-defined tasks. Cited data supports this: Opus only comments on 11% of its tool calls, compared to 65% for Fable. Meanwhile, GPT models remain superior at strictly following complex instructions.

Original post →

More from coding & agent

coding & agent channel →