Mathematics benchmark
HMMT Feb 2026 leaderboard
Harvard-MIT Mathematics Tournament February 2026. Every model the catalog carries a published HMMT Feb 2026 value for, ranked by that value.
A February 2026 HMMT slice used in newer frontier-model math comparisons.
HMMT Feb 2026 ranking
23 models with a published HMMT Feb 2026 value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.
| Rank | Model | Provider | Contest mathematics |
|---|---|---|---|
| 1 | Alibaba | 97.1 | |
| 2 | DeepSeek | 95.2 | |
| 3 | DeepSeek | 94.8 | |
| 4 | Upstage | 93.9 | |
| 5 | Alibaba | 92.9 | |
| 6 | Moonshot AI | 92.7 | |
| 7 | Z.AI | 92.5 | |
| 8 | TMInkling-Small | Thinking Machines Lab | 90.2 |
| 9 | Alibaba | 87.9 | |
| 10 | Alibaba | 87.8 | |
| 11 | Moonshot AI | 87.1 | |
| 12 | InclusionAI | 87.0 | |
| 13 | Z.AI | 86.4 | |
| 14 | Anthropic | 85.3 | |
| 15 | Microsoft | 84.9 | |
| 16 | Alibaba | 84.3 | |
| 17 | Alibaba | 83.6 | |
| 18 | Z.AI | 82.6 | |
| 19 | LG AI Research | 78.4 | |
| 20 | ZYZAYA1-8B | Zyphra | 71.6 |
| 21 | OPMiniCPM5-2B | OpenBMB | 63.8 |
| 22 | Meituan | 40.5 | |
| 23 | OPMiniCPM5-1B | OpenBMB | 25.8 |
Evidence key: Observed
Rows are ordered by the value Qwen3.6 launch benchmarks published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.