The AI arena is free today

Open Superagent

PLawBench

Progress Over Time

Interactive timeline showing model performance evolution on PLawBench

State-of-the-art frontier
Open
Proprietary

PLawBench Leaderboard

1 models
ContextCostLicense
1
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
2.4T1.0M$1.65 / $4.95
Notice missing or incorrect data?
About this benchmark

What is PLawBench?

PLawBench evaluates language models on professional legal knowledge and reasoning tasks.

PLawBench is a text benchmark evaluating models on knowledge, legal, and reasoning tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.

Compare leaders on the best AI for knowledge, best AI for legal and best AI for reasoning leaderboards.

Current leaders

Qwen3.8 Max from Alibaba Cloud / Qwen Team currently leads the PLawBench leaderboard with a score of 0.732 across 1 evaluated AI models.

1Qwen3.8 MaxAlibaba Cloud / Qwen Team73.2%

FAQ

Common questions about the PLawBench benchmark and leaderboard.

What is the PLawBench benchmark?

PLawBench evaluates language models on professional legal knowledge and reasoning tasks.

What is the PLawBench leaderboard?

The PLawBench leaderboard ranks 1 AI models based on their performance on this benchmark. Currently, Qwen3.8 Max by Alibaba Cloud / Qwen Team leads with a score of 0.732. The average score across all models is 0.732.

What is the highest PLawBench score?

The highest PLawBench score is 0.732, achieved by Qwen3.8 Max from Alibaba Cloud / Qwen Team.

How many models are evaluated on PLawBench?

1 models have been evaluated on the PLawBench benchmark, with 0 verified results and 1 self-reported results.

What categories does PLawBench cover?

PLawBench is categorized under knowledge, legal, and reasoning. The benchmark evaluates text models.

What is the best open-source model on PLawBench?

Qwen3.8 Max by Alibaba Cloud / Qwen Team is the top-ranked open-source model on PLawBench, with a score of 0.732 (rank #1).

Which model offers the best value on PLawBench?

Among models scoring within 10% of the leader, Qwen3.8 Max from Alibaba Cloud / Qwen Team is the cheapest, at $1.65 per million input tokens with a score of 0.732.

How recent are the PLawBench leaderboard results?

The PLawBench leaderboard was last updated in August 2026 and currently includes 1 evaluated models.
PLawBench Leaderboard