gpt.college

MMMU-Pro

A harder multimodal academic benchmark derived from MMMU-style tasks.

Last updated: 2026-06-11

What it measures

  • Multimodal academic QA
  • Image-text reasoning
  • Expert subject knowledge

Sourced Ranking

Rank Model Provider Score Source Source Type Verified
1 OpenAI GPT-5.4 OpenAI 81.2 Introducing GPT-5.4 Accessed 2026-06-11 official Yes
2 Gemini 3 Pro Google 81 Gemini 3.1 Pro model page Accessed 2026-06-11 official Yes
3 Gemini 3.1 Pro Google 80.5 Gemini 3.1 Pro model page Accessed 2026-06-11 official Yes
4 OpenAI GPT-5.2 Thinking OpenAI 79.5 Introducing GPT-5.2 Accessed 2026-06-11 official Yes

Limitations

  • New high scores may appear on model release pages before centralized boards update.
  • The arXiv paper is static; leaderboard and Hugging Face sources are better.
  • MMMU official page and HF dataset are selected entry points.

official

MMMU-Pro dataset or code

Use for dataset, code, or implementation details; score freshness depends on the benchmark.