Skip to main content
ModelScale

Coding benchmark

LiveCodeBench Pro leaderboard

Every model the catalog carries a published LiveCodeBench Pro value for, ranked by that value.

CategoryCoding
MeasureCompetitive programming
TasksQuarter-specific contest programming sets
DifficultyHigh-end contest programming

A harder competitive-programming benchmark family built from Codeforces, ICPC, and IOI problems, with quarter-specific public leaderboards and difficulty-aware reporting.

LiveCodeBench Pro ranking

8 models with a published LiveCodeBench Pro value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published LiveCodeBench Pro value
RankModelProviderCompetitive programming
1SASakana Fugu-UltraSakana AI90.8
2SASakana FuguSakana AI87.8
3GPT-5.4OpenAI87.5
4Gemini 3.1 ProGoogle82.9
5Muse SparkMeta80.0
6Grok 4.20xAI74.2
7Claude Opus 4.6Anthropic70.7
8OPMiniCPM5-1BOpenBMB22.7

Evidence key: Observed

Rows are ordered by the value LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards · Coding capability leaderboard