Artificial Analysis breaks down individual evals in Intelligence Index v4.3.2

ArtificialAnlys · x · 2026-09-30

Artificial Analysis published the breakdown of individual evaluations in its Intelligence Index v4.3.2, which combines 10 benchmarks — AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience and AA-LCR v1.1 — across 684 models.

Related event: Artificial Analysis Updates Intelligence Index and Benchmarks GPT-6.1 Sol Variants(2 posts)→

Original post →

More from Models

Models channel →