HMMT Feb 26
Progress Over Time
Interactive timeline showing model performance evolution on HMMT Feb 26
HMMT Feb 26 Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Alibaba Cloud / Qwen Team | — | 1.0M | $1.25 / $3.75 | ||
| 2 | DeepSeek | 1.6T | — | — | ||
| 3 | DeepSeek | 284B | 1.0M | $0.14 / $0.28 | ||
| 4 | Alibaba Cloud / Qwen Team | — | — | — | ||
| 5 | Moonshot AI | 1.0T | 262K | $0.75 / $3.50 | ||
| 6 | Zhipu AI | 753B | 1.0M | $0.95 / $3.00 | ||
| 7 | DeepSeek | 284B | 1.0M | $0.10 / $0.20 | ||
| 8 | Alibaba Cloud / Qwen Team | — | 1.0M | $0.50 / $3.00 | ||
| 9 | Microsoft | 1.0T | — | — | ||
| 10 | Alibaba Cloud / Qwen Team | 28B | 262K | $0.60 / $3.60 | ||
| 11 | Alibaba Cloud / Qwen Team | 35B | — | — | ||
| 12 | Zhipu AI | 754B | 200K | $1.40 / $4.40 |
What is HMMT Feb 26?
HMMT February 2026 is a math competition benchmark based on problems from the Harvard-MIT Mathematics Tournament, testing advanced mathematical problem-solving and reasoning.
HMMT Feb 26 is a text benchmark evaluating models on math and reasoning tasks. LLM Stats tracks 12 models on this benchmark, scored on a 0–1 scale. The current average is 0.9, with the leader at 1.0.
Compare leaders on the best AI for math and best AI for reasoning leaderboards.
Current leaders
Qwen3.7 Max from Alibaba Cloud / Qwen Team currently leads the HMMT Feb 26 leaderboard with a score of 0.971 across 12 evaluated AI models.
FAQ
Common questions about the HMMT Feb 26 benchmark and leaderboard.