MedXpertQA-MM
Progress Over Time
Interactive timeline showing model performance evolution on MedXpertQA-MM
MedXpertQA-MM Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Alibaba Cloud / Qwen Team | — | 1.0M | $0.32 / $1.28 |
What is MedXpertQA-MM?
MedXpertQA-MM is the multimodal subset of MedXpertQA, evaluating expert-level medical question answering grounded in medical images.
MedXpertQA-MM is a multimodal benchmark evaluating models on knowledge, medical, multimodal, and vision tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.
Compare leaders on the best AI for knowledge, best AI for medical, best AI for multimodal and best AI for vision leaderboards.
Current leaders
Qwen3.7-Plus from Alibaba Cloud / Qwen Team currently leads the MedXpertQA-MM leaderboard with a score of 0.710 across 1 evaluated AI models.
FAQ
Common questions about the MedXpertQA-MM benchmark and leaderboard.