Skip to main content
ModelScale

Mathematics benchmark

AIME 2025 leaderboard

American Invitational Mathematics Examination 2025. Every model the catalog carries a published AIME 2025 value for, ranked by that value.

CategoryMathematics
MeasureInteger answers 000-999
Tasks15 problems
DifficultyHigh school olympiad level

The most recent AIME examination, featuring 15 challenging mathematics problems testing olympiad-level mathematical reasoning with integer answers from 000-999.

AIME 2025 ranking

16 models with a published AIME 2025 value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published American Invitational Mathematics Examination 2025 value
RankModelProviderInteger answers 000-999
1MAI-Thinking-1Microsoft97
2Kimi K2.5Moonshot AI96.1
2Kimi K2.5 (Reasoning)Moonshot AI96.1
4GLM-4.7Z.AI95.7
5PMTernary Bonsai 2 27BPrism ML95
6MiMo-V2-FlashXiaomi94.1
7Granite 4.2 30BIBM89.17
8Claude Sonnet 4.5Anthropic87
9Granite 4.2 8BIBM86.67
10OPMiniCPM5-2BOpenBMB86.5
11Exaone 4.0 32BLG AI Research85.3
12Nemotron 3 Nano Omni 30B A3BNVIDIA82.1
13Granite 4.2 3BIBM78.33
14LFM2.5-2.6BLiquidAI51.87
15LFM2.5-8B-A1BLiquidAI42.53
16OPMiniCPM5-1BOpenBMB40.42

Evidence key: Observed

Rows are ordered by the value American Invitational Mathematics Examination published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards