How is Claude Opus 5.5 40% cheaper, 30% faster, and just as good? Reddit theorizes
Witty_County5128 · reddit · 2026-09-28
A Reddit thread dissects Anthropic's Claude Opus 5.5 claims: 40% cheaper to run than Opus 5, 30%+ faster output, and benchmark parity with Fable 5.1 while beating GPT-6 Astra on most evals — getting cheaper, faster, and better simultaneously, which normally means picking two.
The poster notes real developers and companies independently confirm the gains, then speculates on how it's achieved: a fresh pretraining run on cleaner data, post-training/RL making better use of existing compute, distillation from a larger internal model, inference-side optimizations like caching and speculative decoding, or something more fundamental like a looped/recurrent-depth transformer reusing weights. With Anthropic publishing no architecture details, the thread asks ML practitioners to separate realistic engineering from marketing.
More from Models
- "System 2 models built the brain, but System 1 is building the nervous system" — ai · 2026-09-28
- TeleOCR Trends on Hugging Face: A Qwen2.5-VL-Based Chinese Document OCR Model — XingChen-AGI · 2026-09-28
- Kaggle Game Arena: Evaluating LLMs via Head-to-Head Chess, Poker, and Werewolf — kaggle · 2026-09-28
- Perplexity CEO: still using sol 6 for knowledge work — cheap, fast, great compaction — gabriel1 · 2026-09-28
- NerfBench's First Results Find No Nerf: Claude Opus 5.5 Dips Just 0.8% vs Launch — alejandroll10 · 2026-09-28
- Most humans can read this image instantly — most AI vision models can't — JeremyNguyenPhD · 2026-09-28