Claude Opus 5.5 tops ValsAI's RSI Index, first model to beat LM Training reference
scaling01 · x · 2026-09-23
Eval lab ValsAI reports that Claude Opus 5.5 takes #1 on its RSI Index and becomes the first model to beat the published reference on LM Training under its protocol — which ValsAI calls a major step forward for long-horizon agentic work.
More from Models
- GPT-6 Sol and Luna already usable in Codex, early user reports — airesearch12 · 2026-09-23
- GPT-6 Sol scores slightly below GPT-5.6 Sol on DeepSWE, only cheaper — Angaisb_ · 2026-09-23
- Matt Shumer on Opus 5.5: 'feels like a much smarter Opus 4.6' — mattshumer_ · 2026-09-23
- GPT-6 Sol claimed to cost 50% less than GPT-5.6 Sol — cedric_chee · 2026-09-23
- Four frontier models in days: Grok 4.7, Opus 5.5, GPT-6 Sol and Luna — msg · 2026-09-23
- OpenAI reportedly rolling out GPT-6 Sol and GPT-6 Luna on ChatGPT, Codex, and APIs — testingcatalog · 2026-09-23