Skip to main content
ModelScale

korean benchmark

KMMLU leaderboard

Korean Massive Multitask Language Understanding. Every model the catalog carries a published KMMLU value for, ranked by that value.

Categorykorean
MeasureMultiple choice questions
Tasks35,030 questions
DifficultyElementary to professional level in Korean

Evaluates Korean expert-level knowledge across 45 subjects. 20% of questions require Korean cultural context.

KMMLU ranking

2 models with a published KMMLU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Korean Massive Multitask Language Understanding value
RankModelProviderMultiple choice questions
1KAKanana-2 3B InstructKakao43.32
2KAKanana-2 1.3B InstructKakao42.79

Evidence key: Observed

Rows are ordered by the value KMMLU: Measuring Massive Multitask Language Understanding in Korean published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards