GPT-6 Sol Gains 2 Points in Coding Agent Index at Half the Cost
ArtificialAnlys · x · 2026-09-23
Follow-up from Artificial Analysis: in OpenAI's Codex harness, GPT-6 Sol (max) scores 57 in the Coding Agent Index (+2 over GPT-5.6) at roughly half the cost per task, while Luna (max) regresses 2 points. A sub-finding of its full GPT-6 evaluation.
More from Models
- Leaked GPT-6-Sol testing claims half the tokens and 1/5 the time of GPT-5.6-Sol — pvncher · 2026-09-23
- OpenAI blog hints new model is similar size; caching and inference gains double intelligence per dollar — eliebakouch · 2026-09-23
- Dev calls 3D render evals 'mid': rerunning the same prompt beats any model gap — BLUECOW009 · 2026-09-23
- Baseten's head of model training explains why RL extends reasoning but generalization needs domain-specific post-training — baseten · 2026-09-23
- GPT-6 Sol and Claude Opus 5.5 Launch Same Day, Every Runs Head-to-Head Vibe Check — danshipper · 2026-09-23
- Anthropic Opus 5.5 and OpenAI GPT-6 Sol/Luna both promise more capability for less money — Ars Technica AI · 2026-09-23