Hands-on: Comparing Anthropic Opus 4.8 vs. Opus 5 on 25 Tasks
bisonbear · hn · 2026-08-26
The author conducted a comparative test of Anthropic Opus 4.8 vs. Opus 5 across 25 personal tasks. Results show similar final scores, but distinct differences in the reasoning paths and processes used to achieve results, highlighting behavioral changes following the model update.
More from Models
- Zhipu GLM-5.3-Flash: Matches Opus 4.8 at 1/40 the Cost, Powered by Domestic Chips — vista8 · 2026-08-27
- TokenSpeed adds Day-0 support for Qwen 3.8 Flash Next architecture — Alibaba_Qwen · 2026-08-27
- Zhipu GLM-5.3 open weights releasing in 22 hours — Yuchenj_UW · 2026-08-27
- AI models show more creativity when talking to each other than in assistant persona — nabeelqu · 2026-08-27
- OpenRouter leaderboard: Real token consumption data outweighs media hype — sujingshen · 2026-08-27
- Qwen 3.8-Next Released with Detailed Technical Report on Architecture — nrehiew_ · 2026-08-27