Skip to main content
ModelScale

Knowledge benchmark

MMMLU leaderboard

Every model the catalog carries a published MMMLU value for, ranked by that value.

CategoryKnowledge
MeasureExact match
TasksMultilingual academic QA
DifficultyBroad multilingual knowledge
Published byMMMLU

A multilingual MMLU-style benchmark reported in provider evaluation tables.

MMMLU ranking

5 models with a published MMMLU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published MMMLU value
RankModelProviderExact match
1INInterfaze BetaInterfaze90.9
2Qwen3.7 MaxAlibaba90.3
3Qwen3.7 PlusAlibaba89.0
4K-EXAONE 2.0LG AI Research86.6
5Gemma 4 12BGoogle83.4

Evidence key: Observed

Rows are ordered by the value MMMLU published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards · Knowledge capability leaderboard