Repeated tests of Opus 5 suggest lower token use but shallower reasoning
Physical_Concert_625 · reddit · 2026-07-25
The author says repeated testing of Opus 5 over the past day suggests it is being overhyped:
- They claim the model does use fewer tokens, but that comes with weaker reasoning, shallower thinking, and worse decisions.
- In their view, the performance is nowhere near the claims being made compared with other LLMs.
- They ask others to share their own experience testing the model.
Related event: Claude Opus 5 Early Tests: Strong Coding but Disrupts Workflows(14 posts)→
More from Models
- GPT-5.6 Sol looked cheaper across five test cases, and that changes agent margins — PrajwalTomar_ · 2026-07-25
- Claude Opus 5 looks like a major jump over Opus 4.8 on max effort — legit_api · 2026-07-25
- CNBC says distillation is now the fight over who can train from whom — shashib · 2026-07-25
- Open-source momentum builds as Kimi K3 and other frontier labs ship new models — demian_ai · 2026-07-25
- Joke post says AI can finally count to 100 with GPT-5.6 Sol Extra High — BananaIsles · 2026-07-25
- Claude Opus 5 builds a full browser FPS game in one 290 KB HTML file — prasenx · 2026-07-25