The AI arena is free today

Open Superagent

Artificial Analysis

Progress Over Time

Interactive timeline showing model performance evolution on Artificial Analysis

State-of-the-art frontier
Open
Proprietary

Artificial Analysis Leaderboard

8 models
ContextCostLicense
1—1.1M$5.00 / $30.00
2320B1.0M$0.15 / $0.50
3—1.0M$0.75 / $3.75
4—1.1M$2.00 / $12.00
5—500K$2.00 / $6.00
6—1.1M$0.20 / $1.20
7—205K$0.30 / $1.20
8
Thinking Machines Lab
Thinking Machines Lab
276B524K$0.30 / $1.20
Notice missing or incorrect data?
About this benchmark

What is Artificial Analysis?

Artificial Analysis benchmark evaluates AI models across quality, speed, and pricing dimensions, providing a composite assessment of model capabilities for real-world usage.

Artificial Analysis is a text benchmark evaluating models on general tasks. LLM Stats tracks 8 models on this benchmark, scored on a 0–1 scale. The current average is 0.5, with the leader at 0.6.

Compare leaders on the best AI for general leaderboards.

Current leaders

GPT-5.6 Sol from OpenAI currently leads the Artificial Analysis leaderboard with a score of 0.590 across 8 evaluated AI models.

1GPT-5.6 SolOpenAI59.0%
2GLM-5.3-FlashZhipu AI57.0%
3Gemini 3.7 FlashGoogle56.0%

FAQ

Common questions about the Artificial Analysis benchmark and leaderboard.

What is the Artificial Analysis benchmark?

Artificial Analysis benchmark evaluates AI models across quality, speed, and pricing dimensions, providing a composite assessment of model capabilities for real-world usage.

What is the Artificial Analysis leaderboard?

The Artificial Analysis leaderboard ranks 8 AI models based on their performance on this benchmark. Currently, GPT-5.6 Sol by OpenAI leads with a score of 0.590. The average score across all models is 0.528.

What is the highest Artificial Analysis score?

The highest Artificial Analysis score is 0.590, achieved by GPT-5.6 Sol from OpenAI.

How many models are evaluated on Artificial Analysis?

8 models have been evaluated on the Artificial Analysis benchmark, with 0 verified results and 4 self-reported results.

What categories does Artificial Analysis cover?

Artificial Analysis is categorized under general. The benchmark evaluates text models.

What is the best open-source model on Artificial Analysis?

GLM-5.3-Flash by Zhipu AI is the top-ranked open-source model on Artificial Analysis, with a score of 0.570 (rank #2).

Which model offers the best value on Artificial Analysis?

Among models scoring within 10% of the leader, GLM-5.3-Flash from Zhipu AI is the cheapest, at $0.15 per million input tokens with a score of 0.570.

How recent are the Artificial Analysis leaderboard results?

The Artificial Analysis leaderboard was last updated in October 2026 and currently includes 8 evaluated models.