QwenWebBench
Progress Over Time
Interactive timeline showing model performance evolution on QwenWebBench
QwenWebBench Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Alibaba Cloud / Qwen Team | — | 1.0M | $1.25 / $3.75 | ||
| 2 | Alibaba Cloud / Qwen Team | 28B | 262K | $0.60 / $3.60 |
What is QwenWebBench?
QwenWebBench is an internal front-end code generation benchmark by Qwen. It is bilingual (EN/CN) and spans 7 categories (Web Design, Web Apps, Games, SVG, Data Visualization, Animation, and 3D), using auto-render plus a multimodal judge for code and visual correctness. Scores are reported as BT/Elo ratings.
QwenWebBench is a multimodal benchmark evaluating models on multimodal, agents, and code tasks. LLM Stats tracks 2 models on this benchmark, scored on a 0–2000 scale. The current average is 1527.5, with the leader at 1568.0.
Compare leaders on the best AI for multimodal, best AI for agents and best AI for code leaderboards.
Current leaders
Qwen3.7 Max from Alibaba Cloud / Qwen Team currently leads the QwenWebBench leaderboard with a score of 1568.000 across 2 evaluated AI models.
FAQ
Common questions about the QwenWebBench benchmark and leaderboard.