Someone ran tests across all Claude models and published the results
repligate · x · 2026-09-25
Twitter user celestepoasts shared test results across all Claude models via an external link, reposted by repligate. The actual scores live at the linked page.
More from Models
- After RL training, calling LLMs 'language predictors' is no longer accurate, researcher argues — morqon · 2026-09-25
- Anthropic and OpenAI swap playbooks: generous usage vs. user-hostile limits — OwariDa · 2026-09-25
- Dev claims Codex is 10x less token-efficient than Claude Code: $20 buys one day vs one week — DimitrisPapail · 2026-09-25
- Passed-around take: you're better off treating LLMs as brute-force tools — burny_tech · 2026-09-25
- Ego promo video shows brutal model comparison as Claude Opus 5.5 impresses — vista8 · 2026-09-25
- Opus 5.5 takes #1 on CADArena at 0.750, one-shots 3D animation — hudzah · 2026-09-25