Artificial Analysis · MMMU Pro
The Frontier
Best score over time · one chart, every benchmark
Full rankings
33 models tested · sorted by score
| # | Model | Score |
|---|---|---|
| 1 | 87.7 | |
| 2 | 86.9 | |
| 3 | 86.0 | |
| 4 | 85.6 | |
| 5 | 82.8 | |
| 6 | 82.4 | |
| 7 | 80.7 | |
| 8 | 80.5 | |
| 9 | 80.5 | |
| 10 | 79.7 | |
| 11 | 79.0 | |
| 12 | 78.6 | |
| 13 | 78.5 | |
| 14 | 77.0 | |
| 15 | 76.3 | |
| 16 | 75.0 | |
| 17 | 73.4 | |
| 18 | 73.1 | |
| 19 | 70.3 | |
| 20 | 70.1 | |
| 21 | 69.2 | |
| 22 | 66.8 | |
| 23 | 65.4 | |
| 24 | 64.9 | |
| 25 | 63.2 | |
| 26 | 62.1 | |
| 27 | 56.8 | |
| 28 | 52.9 | |
| 29 | 52.7 | |
| 30 | 43.0 | |
| 31 | 37.5 | |
| 32 | 25.8 | |
| 33 | 14.5 |
Score distribution
Where models cluster
Correlated benchmarks
Pearson r · original research
Benchmarks that track with Artificial Analysis · MMMU Pro
Pearson correlation across models scored on both benchmarks. Closer to 1 = strongly predictive.
Frequently asked
About Artificial Analysis · MMMU Pro
What does Artificial Analysis · MMMU Pro measure?
Artificial Analysis · MMMU Pro is a knowledge benchmark in the BenchGecko catalog. 33 AI models have been tested on it. Scores range from 14.5 to 87.7 out of 100.
Which model leads on Artificial Analysis · MMMU Pro?
Claude Opus 5.5 from Anthropic leads Artificial Analysis · MMMU Pro with a score of 87.7. The median score across 33 tested models is 73.4.
Is Artificial Analysis · MMMU Pro saturated?
No · the top score is 87.7 out of 100 (88%). There is still meaningful room for improvement on Artificial Analysis · MMMU Pro.
Does Artificial Analysis · MMMU Pro predict performance on other benchmarks?
Yes · Artificial Analysis · MMMU Pro scores correlate 0.99 with HLE across 5 shared models. Models that do well on Artificial Analysis · MMMU Pro tend to do well on HLE.
How often is Artificial Analysis · MMMU Pro data refreshed?
BenchGecko pulls updates daily. New model scores on Artificial Analysis · MMMU Pro appear as soon as they are published by Epoch AI or the model provider.
- Category
- Knowledge
- Max score
- 100
- Models
- 33
- Updated
- 2026-09-29
Top on Artificial Analysis · MMMU Pro
Claude Opus 5.5 · 87.7GPT-6 Astra · 86.9GPT-6.1 Sol · 86.0Gemini 3.8 Flash · 85.6Qwen3.8 Max (0902) · 82.8More knowledge benchmarks
Same category · related evaluations