MMMU-Pro (Vals)
Vals AI's own run of MMMU-Pro. Run by Vals AI.
As of 19 Sept 2026, Claude Fable 5.1 leads MMMU-Pro (Vals) on BenchLeader with 90.6%, ahead of Claude Opus 5 at 89.9%, across 90 model configurations with a published result.
- Published by
- Vals AI
- Category
- Multimodal
- Index weight
- Reference only
- Models
- 90
- Data as of
- 19 Sept 2026
Vals AI (vals.ai).
What the test looks like
Vals AI's run of the harder ten-option MMMU-Pro multimodal exam.
How it is scored
Accuracy, run by Vals AI.
What to keep in mind
Shown next to the Artificial Analysis run and the official board.
- 1Claude Fable 5.190.6%
- 2Claude Opus 589.9%
- 3Claude Fable 589.3%
- 4Gemini 3.8 Flash (high)89.1%
- 5Gemini 3.7 Flash (high)89.0%
- 6GPT-5.6 Sol (max)88.8%
- 7Gemini 3.6 Flash (high)88.4%
- 8GPT-5.5 (xhigh)88.3%
- 9Gemini 3.5 Flash (high)88.3%
- 10Gemini 3.1 Pro (high)88.2%
- 11Kimi K388.2%
- 12Qwen3 8 (max)88.0%
- 13Gemini 3 Flash (high)87.6%
- 14GPT-5.4 (xhigh)87.5%
- 15Gemini 3 Pro (high)87.5%
90 of 90
| # | ||||
|---|---|---|---|---|
| 1 | 90.6% | 64.5 | 2026-09-01 | |
| 2 | 89.9% | 66.7 | 2026-09-01 | |
| 3 | 89.3% | 68.3 | 2026-09-01 | |
| 4 | 89.1% | 64.5 | 2026-09-01 | |
| 5 | 89.0% | 64.6 | 2026-09-01 | |
| 6 | 88.8% | 68.8 | 2026-09-01 | |
| 7 | 88.4% | 61.7 | 2026-09-01 | |
| 8 | 88.3% | 67.6 | 2026-09-01 | |
| 9 | 88.3% | 63.6 | 2026-09-01 | |
| 10 | 88.2% | 60.7 | 2026-09-01 | |
| 11 | 88.2% | 64.2 | 2026-09-01 | |
| 12 | 88.0% | 66.5 | 2026-09-01 | |
| 13 | 87.6% | 58.7 | 2026-09-01 | |
| 14 | 87.5% | 65.4 | 2026-09-01 | |
| 15 | 87.5% | 61.1 | 2026-09-01 | |
| 16 | 87.4% | 65.7 | 2026-09-01 | |
| 17 | 86.7% | 62.0 | 2026-09-01 | |
| 18 | 86.6% | 61.9 | 2026-09-01 | |
| 19 | 86.6% | 61.8 | 2026-09-01 | |
| 20 | 86.5% | 65.0 | 2026-09-01 | |
| 21 | 86.3% | 60.5 | 2026-09-01 | |
| 22 | 86.1% | 64.1 | 2026-09-01 | |
| 23 | 86.0% | 55.3 | 2026-09-01 | |
| 24 | 85.5% | 64.5 | 2026-09-01 | |
| 25 | 85.0% | 60.2 | 2026-09-01 | |
| 26 | 84.3% | 58.5 | 2026-09-01 | |
| 27 | 84.2% | 58.5 | 2026-09-01 | |
| 28 | 83.9% | 59.2 | 2026-09-01 | |
| 29 | 83.9% | 62.5 | 2026-09-01 | |
| 30 | 83.6% | 59.0 | 2026-09-01 | |
| 31 | 83.6% | 44.3 | 2026-09-01 | |
| 32 | 83.5% | 59.9 | 2026-09-01 | |
| 33 | 83.2% | 58.7 | 2026-09-01 | |
| 34 | 83.1% | 50.5 | 2026-09-01 | |
| 35 | 83.0% | 57.2 | 2026-09-01 | |
| 36 | 83.0% | 60.8 | 2026-09-01 | |
| 37 | 82.5% | 49.1 | 2026-09-01 | |
| 38 | 81.9% | 52.5 | 2026-09-01 | |
| 39 | 81.5% | 58.6 | 2026-09-01 | |
| 40 | 81.3% | 54.5 | 2026-09-01 | |
| 41 | 81.2% | 57.5 | 2026-09-01 | |
| 42 | 81.1% | 58.2 | 2026-09-01 | |
| 43 | 80.8% | 51.6 | 2026-09-01 | |
| 44 | 80.4% | 53.0 | 2026-09-01 | |
| 45 | 79.7% | 50.8 | 2026-09-01 | |
| 46 | 79.5% | 52.1 | 2026-09-01 | |
| 47 | 79.3% | 53.4 | 2026-09-01 | |
| 48 | 79.3% | 55.7 | 2026-09-01 | |
| 49 | 78.9% | 55.3 | 2026-09-01 | |
| 50 | 77.5% | 57.2 | 2026-09-01 | |
| 51 | 77.4% | 44.5 | 2026-09-01 | |
| 52 | 76.3% | 58.5 | 2026-09-01 | |
| 53 | 75.4% | 48.3 | 2026-09-01 | |
| 54 | 75.1% | 51.7 | 2026-09-01 | |
| 55 | 74.9% | 54.0 | 2026-09-01 | |
| 56 | 73.7% | 51.9 | 2026-09-01 | |
| 57 | 73.6% | 49.1 | 2026-09-01 | |
| 58 | 73.3% | 52.9 | 2026-09-01 | |
| 59 | 72.8% | 52.6 | 2026-09-01 | |
| 60 | 72.7% | 56.0 | 2026-09-01 |
Cite as: BenchLeader, “MMMU-Pro (Vals) leaderboard”, https://www.benchleader.com/benchmarks/vals_mmmu_pro, data as of 19 Sept 2026.
MMMU-Pro (Vals): questions
- What does MMMU-Pro (Vals) measure?
- Vals AI's run of the harder ten-option MMMU-Pro multimodal exam. Scores are reported in percent of tasks solved; higher is better.
- Which AI model leads MMMU-Pro (Vals)?
- Claude Fable 5.1 leads MMMU-Pro (Vals) with 90.6% as of 19 Sept 2026, ahead of Claude Opus 5 at 89.9%.
- How many models have MMMU-Pro (Vals) results?
- 90 model configurations have a MMMU-Pro (Vals) result on BenchLeader, all taken from Vals AI.
- Who runs MMMU-Pro (Vals) and how often is it updated?
- MMMU-Pro (Vals) is published by Vals AI. BenchLeader re-reads the published results every morning and records the date each result was published.
- Does MMMU-Pro (Vals) count toward the BenchLeader Index?
- No. MMMU-Pro (Vals) is shown for reference but left out of the composite index.