The AI arena is free today

Open Superagent

PathMCQA

Paper

Progress Over Time

Interactive timeline showing model performance evolution on PathMCQA

State-of-the-art frontier
Open
Proprietary

PathMCQA Leaderboard

1 models
ContextCostLicense
14B
Notice missing or incorrect data?
About this benchmark

What is PathMCQA?

PathMMU is a massive multimodal expert-level benchmark for understanding and reasoning in pathology, containing 33,428 multimodal multi-choice questions and 24,067 images validated by seven pathologists. It evaluates Large Multimodal Models (LMMs) performance on pathology tasks, with the top-performing model GPT-4V achieving only 49.8% zero-shot performance compared to 71.8% for human pathologists.

PathMCQA is a multimodal benchmark evaluating models on multimodal, reasoning, healthcare, and vision tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.

Compare leaders on the best AI for multimodal, best AI for reasoning, best AI for healthcare and best AI for vision leaderboards.

Current leaders

MedGemma 4B IT from Google currently leads the PathMCQA leaderboard with a score of 0.698 across 1 evaluated AI models.

1MedGemma 4B ITGoogle69.8%

FAQ

Common questions about the PathMCQA benchmark and leaderboard.

What is the PathMCQA benchmark?

PathMMU is a massive multimodal expert-level benchmark for understanding and reasoning in pathology, containing 33,428 multimodal multi-choice questions and 24,067 images validated by seven pathologists. It evaluates Large Multimodal Models (LMMs) performance on pathology tasks, with the top-performing model GPT-4V achieving only 49.8% zero-shot performance compared to 71.8% for human pathologists.

What is the PathMCQA leaderboard?

The PathMCQA leaderboard ranks 1 AI models based on their performance on this benchmark. Currently, MedGemma 4B IT by Google leads with a score of 0.698. The average score across all models is 0.698.

What is the highest PathMCQA score?

The highest PathMCQA score is 0.698, achieved by MedGemma 4B IT from Google.

How many models are evaluated on PathMCQA?

1 models have been evaluated on the PathMCQA benchmark, with 0 verified results and 1 self-reported results.

Where can I find the PathMCQA paper?

The PathMCQA paper is available at https://arxiv.org/abs/2401.16355. The paper details the methodology, dataset construction, and evaluation criteria.

What categories does PathMCQA cover?

PathMCQA is categorized under multimodal, reasoning, healthcare, and vision. The benchmark evaluates multimodal models.

What is the best open-source model on PathMCQA?

MedGemma 4B IT by Google is the top-ranked open-source model on PathMCQA, with a score of 0.698 (rank #1).

How recent are the PathMCQA leaderboard results?

The PathMCQA leaderboard was last updated in August 2026 and currently includes 1 evaluated models.