IPhO 2025
Progress Over Time
Interactive timeline showing model performance evolution on IPhO 2025
IPhO 2025 Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Meta | — | — | — | ||
| 2 | ByteDance | — | — | — |
What is IPhO 2025?
International Physics Olympiad 2025 (theory) comprises all 3 theory problems from the official 2025 IPhO competition. Results are based on blinded human evaluation with guidelines based on the official competition scoring, validated by domain experts.
IPhO 2025 is a text benchmark evaluating models on physics and reasoning tasks. LLM Stats tracks 2 models on this benchmark, scored on a 0–1 scale. The current average is 0.8, with the leader at 0.8.
Compare leaders on the best AI for physics and best AI for reasoning leaderboards.
Current leaders
Muse Spark from Meta currently leads the IPhO 2025 leaderboard with a score of 0.826 across 2 evaluated AI models.
FAQ
Common questions about the IPhO 2025 benchmark and leaderboard.