Reddit says a Gemini 3.5 checkpoint beat Claude Opus 5 Max Thinking in a test
Last_Conclusion_8984 · reddit · 2026-07-27
A Reddit post claims a Gemini 3.5 checkpoint beat Claude Opus 5 Max Thinking in a test.
The attached image shows an Arena-style head-to-head evaluation page with Gemini 3.6 Flash facing Claude Opus 5 Max, suggesting the claim is about a benchmark or internal test result rather than a product announcement.
No additional methodology or score details are provided in the post itself, so the value here is mainly the reported result and the comparison between the two models.
Related event: Rumors: New Gemini 3.5 Checkpoint Impresses and Beats Opus 5 Max(5 posts)→
More from Models
- Writer says ChatGPT still trails Claude in creative writing after a month of testing — NeonXEExperiment · 2026-07-27
- Anthropic publishes an Opus 5 prompt guide built around subtraction, not addition — xiaohu · 2026-07-27
- Google says 86% of Gemini chats in a 14.65M-conversation study were non-work related — xiaohu · 2026-07-27
- Claude Opus 5 tops OSWorld with 70.6% partial-credit, about 30% strict success — ysu_nlp · 2026-07-27
- Reddit users compare Pro and Ultra for deep research and multi-agent R&D work — curiousinquirer007 · 2026-07-27
- Which local models do people still use months after the hype fades? — b111ue · 2026-07-27