Benchmark profile
Massive Multi-discipline Multimodal Understanding Pro (MMMU-Pro)
A harder multimodal benchmark for frontier models that combines text with images, diagrams, charts, and academic visual reasoning tasks.
Data verifiedTop models on MMMU-Pro — July 23, 2026
As of July 23, 2026, GPT-5.4 Pro leads the MMMU-Pro leaderboard with 94% , followed by Gemini 3.1 Pro (83.9%) and Gemini 3.5 Flash (83.6%).
GPT-5.4 Pro
OpenAI
gpt-5-4-pro
Gemini 3.1 Pro
gemini-3-1-pro
Gemini 3.5 Flash
gemini-3-5-flash
Leaderboard (34 models)
ScoreAccording to BenchLM.ai, GPT-5.4 Pro leads the MMMU-Pro benchmark with a score of 94%, followed by Gemini 3.1 Pro (83.9%) and Gemini 3.5 Flash (83.6%). The scores show moderate spread, with meaningful differences between the top tier and mid-tier models.
34 models have been evaluated on MMMU-Pro. The benchmark falls in the Multimodal & Grounded category. This category carries a 12% weight in BenchLM.ai's overall scoring system. Within that category, MMMU-Pro contributes 45% of the category score, so strong performance here directly affects a model's overall ranking.
About MMMU-Pro
Year
2024
Tasks
Multimodal academic reasoning
Format
Image + text question answering
Difficulty
Frontier multimodal
MMMU-Pro extends the original MMMU setup with more difficult multimodal questions and stronger separation at the top end of the model market.
BenchLM freshness & provenance
Version
MMMU-Pro 2024
Refresh cadence
Annual
Staleness state
Refreshing
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
FAQ
What does MMMU-Pro measure?
A harder multimodal benchmark for frontier models that combines text with images, diagrams, charts, and academic visual reasoning tasks.
Which model scores highest on MMMU-Pro?
GPT-5.4 Pro by OpenAI currently leads with a score of 94% on MMMU-Pro.
How many models are evaluated on MMMU-Pro?
34 AI models have been evaluated on MMMU-Pro on BenchLM.
Compare Top Models on MMMU-Pro
Choose a model with this week’s evidence
Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.
One email each week. Unsubscribe anytime.