The AI arena is free today

Open Superagent

HMMT Feb 26

Progress Over Time

Interactive timeline showing model performance evolution on HMMT Feb 26

State-of-the-art frontier
Open
Proprietary

HMMT Feb 26 Leaderboard

12 models
ContextCostLicense
1
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
1.0M$1.25 / $3.75
21.6T
3284B1.0M$0.14 / $0.28
4
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
5
Moonshot AI
Moonshot AI
1.0T262K$0.75 / $3.50
6
Zhipu AI
Zhipu AI
753B1.0M$0.95 / $3.00
7284B1.0M$0.10 / $0.20
8
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
1.0M$0.50 / $3.00
91.0T
10
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
28B262K$0.60 / $3.60
11
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
35B
12
Zhipu AI
Zhipu AI
754B200K$1.40 / $4.40
Notice missing or incorrect data?
About this benchmark

What is HMMT Feb 26?

HMMT February 2026 is a math competition benchmark based on problems from the Harvard-MIT Mathematics Tournament, testing advanced mathematical problem-solving and reasoning.

HMMT Feb 26 is a text benchmark evaluating models on math and reasoning tasks. LLM Stats tracks 12 models on this benchmark, scored on a 0–1 scale. The current average is 0.9, with the leader at 1.0.

Compare leaders on the best AI for math and best AI for reasoning leaderboards.

Current leaders

Qwen3.7 Max from Alibaba Cloud / Qwen Team currently leads the HMMT Feb 26 leaderboard with a score of 0.971 across 12 evaluated AI models.

1Qwen3.7 MaxAlibaba Cloud / Qwen Team97.1%
2DeepSeek-V4-Pro-MaxDeepSeek95.2%
3DeepSeek-V4-Flash-MaxDeepSeek94.8%

FAQ

Common questions about the HMMT Feb 26 benchmark and leaderboard.

What is the HMMT Feb 26 benchmark?

HMMT February 2026 is a math competition benchmark based on problems from the Harvard-MIT Mathematics Tournament, testing advanced mathematical problem-solving and reasoning.

What is the HMMT Feb 26 leaderboard?

The HMMT Feb 26 leaderboard ranks 12 AI models based on their performance on this benchmark. Currently, Qwen3.7 Max by Alibaba Cloud / Qwen Team leads with a score of 0.971. The average score across all models is 0.900.

What is the highest HMMT Feb 26 score?

The highest HMMT Feb 26 score is 0.971, achieved by Qwen3.7 Max from Alibaba Cloud / Qwen Team.

How many models are evaluated on HMMT Feb 26?

12 models have been evaluated on the HMMT Feb 26 benchmark, with 0 verified results and 12 self-reported results.

What categories does HMMT Feb 26 cover?

HMMT Feb 26 is categorized under math and reasoning. The benchmark evaluates text models.

What is the best open-source model on HMMT Feb 26?

DeepSeek-V4-Pro-Max by DeepSeek is the top-ranked open-source model on HMMT Feb 26, with a score of 0.952 (rank #2).

Which model offers the best value on HMMT Feb 26?

Among models scoring within 10% of the leader, DeepSeek-V4-Flash-0423 from DeepSeek is the cheapest, at $0.10 per million input tokens with a score of 0.919.

How recent are the HMMT Feb 26 leaderboard results?

The HMMT Feb 26 leaderboard was last updated in August 2026 and currently includes 12 evaluated models.