Skip to main content
ModelScale

Mathematics benchmark

AA AIME 2025 leaderboard

Artificial Analysis AIME 2025. Every model the catalog carries a published AA AIME 2025 value for, ranked by that value.

CategoryMathematics
MeasureAccuracy
Tasks30 AIME 2025 problems
DifficultyOlympiad mathematics

An independently evaluated AIME 2025 result from Artificial Analysis.

AA AIME 2025 ranking

2 models with a published AA AIME 2025 value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Artificial Analysis AIME 2025 value
RankModelProviderAccuracy
1GPT-5.2OpenAI99.0
2GPT-OSS 120BOpenAI93.4

Evidence key: Observed

Rows are ordered by the value Artificial Analysis AIME 2025 Benchmark Leaderboard published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards