Skip to main content
ModelScale

Multilingual benchmark

MILU leaderboard

Multi-task Indic Language Understanding Benchmark. Every model the catalog carries a published MILU value for, ranked by that value.

CategoryMultilingual
MeasureAverage accuracy
TasksKnowledge tasks across 11 languages
DifficultyMultilingual Indic knowledge

Culturally grounded knowledge comprehension across ten Indic languages and English.

MILU ranking

1 model with a published MILU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Multi-task Indic Language Understanding Benchmark value
RankModelProviderAverage accuracy
1Claude Opus 5Anthropic92.1

Evidence key: Observed

Rows are ordered by the value MILU: A Multi-task Indic language understanding benchmark published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards