Composite indices
Artificial Analysis Intelligence Index v4.3.2
Artificial Analysis composite across AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, and AA-LCR v1.1. Keep separate from prior index versions because its evaluation components changed.
How models are tested
Scores are published numbers from the sources below. They are only comparable when the evaluation version, tools, agent harness, and sampling setup match. Missing scores are omitted rather than treated as zero.
- Category
- Composite indices
- Metric
- Index (0-100)
- Direction
- Higher is better
- Coverage
- 8 / 318Models in this catalog with a published score
- Updated
- Catalog snapshot date
Scores
- 1Claude Opus 5.558.0
Anthropicclosed
- 2Claude Sonnet 5.556.0
Anthropicclosed
- 3GPT-6 Astra53.0
OpenAIclosed
- 4Claude Fable 5.153.0
Anthropicclosed
- 5GPT-6.1 Sol52.0
OpenAIclosed
- 6GPT-6 Sol48.0
OpenAIclosed
- 7Grok 4.746.0
xAIclosed
- 8GPT-6 Luna37.0
OpenAIclosed