Claude Opus 5.5 Benchmark Scores Circulating on Reddit
FalconsArentReal · reddit · 2026-09-23
A Reddit user shared a screenshot of alleged Claude Opus 5.5 benchmark scores, showing results across multiple evaluation suites — worth comparing against Anthropic's official numbers.
More from Models
- OpenCompare helps you pick the best open model by cost and quality — nutlope · 2026-09-23
- Leak claims GPT-6-sol is up to 6x faster than astra in grey testing; agent group-chat teased — basedjensen · 2026-09-23
- Anthropic quietly adds banked resets for Claude usage limits — airesearch12 · 2026-09-23
- Anthropic report argues AI R&D evals are saturated and uninformative — dfrsrchtwts · 2026-09-23
- Reddit rumor points to imminent AI release wave: GPT-6, Opus 5.5, Fable 5.5 — 141_1337 · 2026-09-23
- Dev calls out 626 passing tests as junk, suspects RL side effect — dyn___ · 2026-09-23