Artificial Analysis updates: GPT-6.1 Sol tops GPT-6 Sol in 7 days, Claude Sonnet 5.5 hits #2
ArtificialAnlys · x · 2026-09-30
Artificial Analysis' linked page details the evaluations behind its intelligence-cost Pareto frontier analysis:
- GPT-6.1 Sol replaced GPT-6 Sol after just 7 days, with near-Astra intelligence; all effort tiers are now evaluated.
- Claude Sonnet 5.5 (Adaptive Reasoning, multiple effort settings) reached #2 on the Intelligence Index.
- A new Cyber Index addresses cyber defense capability evaluation.
- Index v4.3 replaced 𝜏³-Banking with AutomationBench-AA and upgraded Terminal-Bench to 4.0.
- New Optima custom benchmark builder and model recommender also launched.
Related event: GPT-6.1 Sol Launches with Astra-level Intelligence at One-fifth the Price(32 posts)→
More from Models
- Dev praises Opus 5.5 status messages for reporting progress like an expert engineer — sytelus · 2026-10-01
- Untrained d1 model plays Smash Melee at impressive level, researcher stunned — maximelabonne · 2026-10-01
- GPT-6 Astra prefers CDT over EDT, unlike Anthropic models, suggests decision-theory sycophancy gap — dfrsrchtwts · 2026-10-01
- OpenAI makes Luna free for subscribers; dev says gap to bigger models is now narrow — dosco · 2026-10-01
- 8B model trained on distributed gaming GPUs for $6,500 runs on phone CPU at ~60 tok/s — markjeffrey · 2026-10-01
- Dev hails Claude Opus 5.5 as the best coding experience since Opus 4.5 — jarrodwatts · 2026-10-01