Artificial Analysis Intelligence Index v4.3 raises private-question weight to 45%
ArtificialAnlys · x · 2026-09-08
Artificial Analysis published the methodology for Intelligence Index v4.3:
- Category weights unchanged: Agents 30%, Coding 20%, General 30%, Scientific Reasoning 20%
- Terminal-Bench 4.0 keeps the same weighting as 2.1, with AutomationBench-AA replacing 𝜏³-Banking at its 5% weight
- Evaluations with private questions/answers now account for 45% of the weighting, up from 40% in v4.2, reducing benchmark contamination risk
Full results and methodology are available on the Artificial Analysis site.
More from Models
- Apodex 1.1 Agent Team hits 63.3% pass rate on FrontierScience-Research, paper lands on Papers with Code — NielsRogge · 2026-09-08
- Gemini Pro runs research task for nearly 5 hours with barely any progress — teortaxesTex · 2026-09-08
- "Infinite Wikipedia That Talks Back": Using Smart LLMs to Get Deliberately Lost — signulll · 2026-09-08
- Astra Tries to Reproduce EUV Scattering Paper Computationally, Fails — teortaxesTex · 2026-09-08
- Command Code ships Muse Spark 1.3 Max with new reasoning effort parameter — alexandr_wang · 2026-09-08
- Jensen Huang confirms GPT-6 Astra trained on 100K+ Grace Blackwell NVL72 systems — rohanpaul_ai · 2026-09-08