Claude Opus 5 Hands-on: Impressive Single-Prompt Generation, but Lags Behind Fable in Complex Tasks

Early adopters recently shared hands-on experiences with Claude Opus 5, widely praising its exceptional common sense and instruction following, noting it even exceeds expectations in non-thinking mode.

Confirmed

Based on collective user feedback, Opus 5 demonstrates the following traits in real-world tasks:

1. **Clear division of labor**: User @dr_cintas outlined a usage strategy, recommending Sonnet 5 for daily coding, drafting, and quick edits, while reserving Opus 5 for complex agent tasks.

2. **Excellent real-world performance**: @dejavucoder noted a preference for using Opus 5's mid-to-high settings in practice, drastically reducing their use of Sonnet 5.

3. **Strong common sense and instruction following**: Both @dejavucoder and @JasonBotterill stated that Opus 5 better grasps user intent, whereas GPT-5.6 Sol feels relatively rigid in instruction adherence by comparison.

Unconfirmed

The performance comparison between **non-thinking mode Opus 5** and **thinking mode Sonnet 5** is currently based solely on subjective speculation and personal experiences from users like @dejavucoder and @JasonBotterill. @JasonBotterill specifically highlighted that while non-thinking Opus 5 is highly capable, it hasn't yet undergone systematic benchmark testing, and Anthropic has not officially featured this comparison.

Why it matters

These in-depth user tests reveal the real-world potential of Anthropic's new model, particularly the foundational improvements in "common sense." This not only provides a reference for developers choosing models but also reflects the current competitive landscape among top-tier AI models regarding instruction following and flexibility.

2026-07-26 ~ 2026-07-28 · 24 related posts

Full story(20 episodes)→

Primary sources

1 near-duplicate retellings: eyishazyer