Skip to main content
ModelScale

Multimodal & Grounded benchmark

MMMU leaderboard

Massive Multi-discipline Multimodal Understanding. Every model the catalog carries a published MMMU value for, ranked by that value.

CategoryMultimodal & Grounded
MeasureImage + text question answering
TasksMultimodal academic reasoning
DifficultyFrontier multimodal

A broad multimodal reasoning benchmark spanning charts, diagrams, tables, and academic visual question answering.

MMMU ranking

11 models with a published MMMU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Massive Multi-discipline Multimodal Understanding value
RankModelProviderImage + text question answering
1Qwen3.6 PlusAlibaba86.0
2Qwen3.5-122B-A10BAlibaba83.9
3Qwen3.6-27BAlibaba82.9
4Qwen3.5-27BAlibaba82.3
5Qwen3.6-35B-A3BAlibaba81.7
6Qwen3.5-35B-A3BAlibaba81.4
7Command A+Cohere75.1
8Nemotron 3 Nano Omni 30B A3BNVIDIA70.8
9LFM2.5-VL-3BLiquidAI48.4
10ZYZAYA1-VL-8BZyphra46.0
11LFM2.5-VL-450MLiquidAI32.7

Evidence key: Observed

Rows are ordered by the value MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards