Claude Opus 5 Hands-on: Impressive Single-Prompt Generation, but Lags Behind Fable in Complex Tasks
Early adopters recently shared hands-on experiences with Claude Opus 5, widely praising its exceptional common sense and instruction following, noting it even exceeds expectations in non-thinking mode.
Confirmed
Based on collective user feedback, Opus 5 demonstrates the following traits in real-world tasks:
1. **Clear division of labor**: User @dr_cintas outlined a usage strategy, recommending Sonnet 5 for daily coding, drafting, and quick edits, while reserving Opus 5 for complex agent tasks.
2. **Excellent real-world performance**: @dejavucoder noted a preference for using Opus 5's mid-to-high settings in practice, drastically reducing their use of Sonnet 5.
3. **Strong common sense and instruction following**: Both @dejavucoder and @JasonBotterill stated that Opus 5 better grasps user intent, whereas GPT-5.6 Sol feels relatively rigid in instruction adherence by comparison.
Unconfirmed
The performance comparison between **non-thinking mode Opus 5** and **thinking mode Sonnet 5** is currently based solely on subjective speculation and personal experiences from users like @dejavucoder and @JasonBotterill. @JasonBotterill specifically highlighted that while non-thinking Opus 5 is highly capable, it hasn't yet undergone systematic benchmark testing, and Anthropic has not officially featured this comparison.
Why it matters
These in-depth user tests reveal the real-world potential of Anthropic's new model, particularly the foundational improvements in "common sense." This not only provides a reference for developers choosing models but also reflects the current competitive landscape among top-tier AI models regarding instruction following and flexibility.
2026-07-26 ~ 2026-07-28 · 24 related posts
- Episode 1: GPT-5.6 Variants Revealed, Rumored to Launch by July 7(2026-07-03, 8 posts)
- Episode 2: GPT 5.6 Is Opus-Tier, Cheaper and Faster Than Opus 4.8(2026-07-04, 3 posts)
- Episode 3: Rumors Swirl Around Impending Release of OpenAI's GPT-5.6 Series(2026-07-05, 17 posts)
- Episode 4: Unverified Rumor Says GPT-5.6 Found New Math(2026-07-06, 2 posts)
- Episode 5: Musk Announces Grok 4.5 with 1.5T Parameters and Enhanced Coding(2026-07-07, 25 posts)
- Episode 6: Prediction Markets Strongly Price In Grok 4.4 Release(2026-07-07, 2 posts)
- Episode 7: OpenAI Announces GPT-5.6 Sol for Thursday Release Amid Early Tester Reviews(2026-07-07, 58 posts)
- Episode 8: OpenAI Launches Full-Duplex Voice Model GPT-Live(2026-07-07, 44 posts)
- Episode 9: Grok 4.5 Released with Focus on Coding and Low Cost(2026-07-08, 61 posts)
- Episode 10: New ChatGPT Voice Mode Tested: Near-Human Multi-lingual Experience(2026-07-09, 14 posts)
- Episode 11: GPT-5.6 Tested: Major Coding Leap and Direct Rival to Fable 5(2026-07-09, 30 posts)
- Episode 12: xAI Launches Grok 4.5: Coding and Agent Focus to Rival Opus(2026-07-09, 55 posts)
- Episode 13: Grok 4.5 Benchmarks Strong but Faces Data Controversy(2026-07-09, 6 posts)
- Episode 14: Rumors Swirl Over Imminent Releases of Multiple AI Models(2026-07-09, 2 posts)
- Episode 15: Grok 4.5 Receives Widespread Praise for Speed and Coding(2026-07-09, 13 posts)
- Episode 16: Grok 4.5 Praised for Impressive Speed and Performance(2026-07-09, 2 posts)
- Episode 17: Grok 4.5 Outperforms Fable in Coding Speed and Efficiency(2026-07-09, 3 posts)
- Episode 18: Grok 4.5 Released, Ranks 6th on Vals Index(2026-07-09, 2 posts)
- Episode 19: Frontier Model Comparison: GPT-5.6 Praised for Value and Creativity(2026-07-09, 3 posts)
- Episode 20: OpenAI Launches GPT-5.6 Series: Multi-Agent and Cost-Efficiency(2026-07-09, 119 posts)
Primary sources
- Anthropic users say Opus 5 is best for complex agents, while Sonnet 5 fits routine coding — dr_cintas · 2026-07-26
- Fable 5 beats Opus 5 on open-ended coding tasks, says developer after 24 hours — johnlindquist · 2026-07-26
- [source] Opus 5 beats Fable on benchmarks but loses badly in real use, analyst says — petergyang · 2026-07-26
- Claude Opus 5 sparks a wave of wild early builds within 24 hours — Aiden_Tech_Ai · 2026-07-26
- Opus 5 reportedly delivers remarkable 3D results, despite not winning everywhere — TAbrodi · 2026-07-26
- After two days of testing, antirez says Opus 5 closes the gap but still doesn't beat Sol — antirez · 2026-07-26
- User says Claude Opus 5 feels more forgiving and “common-sense” than GPT-5.6 Sol — dejavucoder · 2026-07-26
- Botterill says Opus 5 feels more commonsense than GPT-5.6 Sol on instruction following — JasonBotterill · 2026-07-26
- Jason Botterill says non-thinking Opus 5 may beat Sonnet 5 on some tasks — JasonBotterill · 2026-07-26
- Users say Opus 5 outperforms Sonnet 5 in their real-world workflow — dejavucoder · 2026-07-26
- Opus 5 code demo improves and the source is now public — bdsqlsz · 2026-07-27
- Reddit user says Opus 5 codes better, but is far more pedantic and hard to steer — Veraticus · 2026-07-27
- [source] User says Opus 5 lags Fable on tricky tasks despite still being a strong model — doodlestein · 2026-07-27
- Opus 5 Struggles in Complex Tasks While Fable Shows Deeper Reasoning — doodlestein · 2026-07-27
- Opus may cost more than Fable on hard debugging when retries pile up — doodlestein · 2026-07-27
- Claude Opus 5 is being described as a strong subagent, not a great ideation partner — iskander · 2026-07-27
- [source] Three days after launch, Opus 5 is already driving a wave of demos — eyishazyer · 2026-07-27
- Three days after launch, Opus 5 is already making consultant-grade decks — eyishazyer · 2026-07-27
- Opus 5 recreates Minecraft with block physics, shadows and lighting — eyishazyer · 2026-07-27
- Opus 5 produces a one-shot snowboard simulation with solid physics — eyishazyer · 2026-07-27
- Opus 5 can build a full custom first-person shooter from one prompt — eyishazyer · 2026-07-27
- Opus 5 looks stronger than Opus 4.8, but picks up some Fable quirks — eyishazyer · 2026-07-27
- Users say Opus 5 breaks things, while Opus 4.8 still feels strong — omarsar0 · 2026-07-28
1 near-duplicate retellings: eyishazyer