PerceptionBench
Progress Over Time
Interactive timeline showing model performance evolution on PerceptionBench
PerceptionBench Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Qwen3.8 MaxNew Alibaba Cloud / Qwen Team | 2.4T | — | — | ||
| 2 | Moonshot AI | 2.8T | 1.0M | $3.00 / $15.00 |
What is PerceptionBench?
PerceptionBench is Moonshot AI's internal benchmark for evaluating atomic visual perception capabilities.
PerceptionBench is a multimodal benchmark evaluating models on multimodal, reasoning, and vision tasks. LLM Stats tracks 2 models on this benchmark, scored on a 0–1 scale. The current average is 0.6, with the leader at 0.6.
Compare leaders on the best AI for multimodal, best AI for reasoning and best AI for vision leaderboards.
Current leaders
Qwen3.8 Max from Alibaba Cloud / Qwen Team currently leads the PerceptionBench leaderboard with a score of 0.635 across 2 evaluated AI models.
FAQ
Common questions about the PerceptionBench benchmark and leaderboard.