PathMCQA
Progress Over Time
Interactive timeline showing model performance evolution on PathMCQA
PathMCQA Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Google | 4B | — | — |
What is PathMCQA?
PathMMU is a massive multimodal expert-level benchmark for understanding and reasoning in pathology, containing 33,428 multimodal multi-choice questions and 24,067 images validated by seven pathologists. It evaluates Large Multimodal Models (LMMs) performance on pathology tasks, with the top-performing model GPT-4V achieving only 49.8% zero-shot performance compared to 71.8% for human pathologists.
PathMCQA is a multimodal benchmark evaluating models on multimodal, reasoning, healthcare, and vision tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.
Compare leaders on the best AI for multimodal, best AI for reasoning, best AI for healthcare and best AI for vision leaderboards.
Current leaders
MedGemma 4B IT from Google currently leads the PathMCQA leaderboard with a score of 0.698 across 1 evaluated AI models.
FAQ
Common questions about the PathMCQA benchmark and leaderboard.