korean benchmark
KMMLU leaderboard
Korean Massive Multitask Language Understanding. Every model the catalog carries a published KMMLU value for, ranked by that value.
Evaluates Korean expert-level knowledge across 45 subjects. 20% of questions require Korean cultural context.
KMMLU ranking
2 models with a published KMMLU value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.
| Rank | Model | Provider | Multiple choice questions |
|---|---|---|---|
| 1 | KAKanana-2 3B Instruct | Kakao | 43.32 |
| 2 | KAKanana-2 1.3B Instruct | Kakao | 42.79 |
Evidence key: Observed
Rows are ordered by the value KMMLU: Measuring Massive Multitask Language Understanding in Korean published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.