Reasoning & knowledge
Chartography
Professional chart-understanding benchmark with 100 real-world chart questions across 12 domains; preserve tool access and scoring setup because tool-enabled and no-tool results are distinct.
How models are tested
Scores are published numbers from the sources below. They are only comparable when the evaluation version, tools, agent harness, and sampling setup match. Missing scores are omitted rather than treated as zero.
- Category
- Reasoning & knowledge
- Metric
- Percent
- Direction
- Higher is better
- Coverage
- 4 / 318Models in this catalog with a published score
- Updated
- Catalog snapshot date
Scores
- 1Claude Opus 5.564.4%
Anthropicclosed
- 2Claude Sonnet 5.561.6%
Anthropicclosed
- 3GPT-6 Sol53.6%
OpenAIclosed
- 4Claude Sonnet 515.6%
Anthropicclosed