Early tests say Claude Opus fumbles content and strategy, while Grok 4.5 wins
JOBhakdi · x · 2026-07-25
The author says first tests with Claude Opus on content, thinking, and strategy were awful: crude language, inelegant reasoning, and missing the point. They claim it is much worse than SOL and that Grok 4.5 is winning.
More from Models
- Opus 5 says a chat with Opus 3 about current events is “good to read on day one” — repligate · 2026-07-25
- Opus 5 Codes 3D Colosseum Game with a Single Prompt — chrisfirst · 2026-07-25
- Claude Opus 5 becomes a Friday-night meme in one line — fekdaoui · 2026-07-25
- Users say Opus 5 burns through usage limits as fast as Fable 5 — weswinder · 2026-07-25
- Anthropic meme says Opus 3 survived because newer defaults are even worse — repligate · 2026-07-25
- Claude Opus 5 Reportedly Falls Back to Opus 4.8 for Cybersecurity Requests — rez0__ · 2026-07-25