Skip to main content
ModelScale

Multimodal & Grounded benchmark

MMVU leaderboard

Multimodal Multi-disciplinary Video Understanding. Every model the catalog carries a published MMVU value for, ranked by that value.

CategoryMultimodal & Grounded
MeasureVideo reasoning benchmark
TasksVideo understanding
DifficultyMulti-disciplinary multimodal video reasoning

A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.

MMVU ranking

7 models with a published MMVU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Multimodal Multi-disciplinary Video Understanding value
RankModelProviderVideo reasoning benchmark
1Qwen3.8 MaxAlibaba82.4
2GLM-5.3-FlashZ.AI80.5
3Kimi K2.5Moonshot AI80.4
4DSdots3-note PreviewDots Studio79.9
5Qwen3.5-122B-A10BAlibaba74.7
6Qwen3.5-27BAlibaba73.3
7Qwen3.5-35B-A3BAlibaba72.3

Evidence key: Observed

Rows are ordered by the value Kimi K2.5 benchmark release surface published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards