Artificial Analysis Breaks Down Individual Evals in Intelligence Index v4.3
ArtificialAnlys · x · 2026-09-23
Artificial Analysis shares a breakdown of individual evaluations in its Intelligence Index v4.3 for GPT-6 Sol and Luna, part of its evaluation thread — previously noting GDPval-AA v2.1 regressions driven by shorter deliverables omitting required elements.
More from Models
- Claude Opus 5.5 Debuts: 40% Cheaper, ~30% Faster Than Opus 5, Stronger Agentic Coding — nicolechirps · 2026-09-23
- Fireworks Launches Specialized Intelligence Index; DFS Model Hits 62.2% Vuln Detection Recall — nicolechirps · 2026-09-23
- Sol 6 matches Sol 5.6 quality with half the reasoning time, better Plus limits — Simple-Diver-2192 · 2026-09-23
- Founder: OpenAI should ship its stronger internal model and cut Astra prices 50% — bindureddy · 2026-09-23
- No Single Champion: Model Rankings Shift by Domain, from Legal to Finance to Healthcare — sophiamyang · 2026-09-23
- Quintin Pope asks if Claude shows in-context emergent misalignment — QuintinPope5 · 2026-09-23