The AI arena is free today

Open Superagent
QwenReleased on Dec 15, 2025

Qwen3 Max: Benchmarks, Pricing & Context Window

Qwen3 Max is a language model from Qwen, released in December 2025, with a 256K-token context window, and pricing from $1.20/M input, $0.240/M cached input, $6.00/M output.

Qwen3 Max is the flagship model in the Qwen3 series, designed for maximum performance across all tasks. It excels in reasoning, mathematics, coding, and complex problem-solving while maintaining strong multilingual capabilities. Optimized

Input
Text
Output
Text

Qwen3 Max benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Qwen3 Max across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Qwen3 Max holds up as conversations get longer.

Quality Tracker

Qwen3 Max Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Fri Oct 09 2026
Notice missing or incorrect data?

Qwen3 Max pricing

Providers

Qwen3 Max starts at $0.500 per million input tokens and $5.00 per million output tokens via Novita. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Novita logoNovita
$0.500—$5.00256.0K/131.1K
—
—
/
DeepInfra logoDeepInfra
$1.20$0.240$6.00256.0K/256.0K
—
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

Qwen3 Max context window

Input and output token limits for Qwen3 Max, plus how it ranks on long-context understanding.

InputOutput
256Ktokens
256Ktokens
≈ 385 pages of text
256K
8K128K1M

Try now

huggle
Qwen3 Maxin Huggle

Make it with
Qwen3 Max.

Qwen3 Max

Qwen3 Max latency

Qwen3 Max time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Qwen3 Max examples

Recent arena outputs from Qwen3 Max, picked from the highest-ranked matchups.

Qwen3 Max license

Qwen3 Max is a proprietary model available under its provider's product and API terms, has 1.0T parameters.

License
Proprietary
Hosted access
Parameters
1.0T

Proprietary license - usage restrictions apply

Qwen3 Max resources

Official sources for Qwen3 Max: provider documentation, official playground, paper or system card, source repository.

Qwen3 Max vs other models

The most-compared alternatives to Qwen3 Max are Claude 3.5 Sonnet, LongCat-Flash-Thinking-2601, Nova 2 Pro. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3 Max

Models ranked just above and below Qwen3 Max by LLM Stats score.

 

Claude 3.5 Sonnet

Score pending
 

LongCat-Flash-Thinking-2601

Score pending
 

Nova 2 Pro

Score pending
 

DeepSeek R1 Distill Qwen 32B

Score pending
 

o1-mini

Score pending
 

Qwen3 VL 30B A3B Thinking

Score pending

FAQ

Common questions about Qwen3 Max.

When was Qwen3 Max released?

Qwen3 Max was released on December 15, 2025 by Qwen. This is the official Qwen3 Max release date tracked on LLM Stats.

How much does Qwen3 Max cost?

Qwen3 Max pricing starts at $0.50 per million input tokens and $5.00 per million output tokens via Novita, the lowest price among tracked providers.

How big is Qwen3 Max?

Qwen3 Max has 1000 billion parameters. It was trained on 36.0 trillion tokens.

Who created Qwen3 Max?

Qwen3 Max was created by Qwen.

What is the license for Qwen3 Max?

Qwen3 Max is released under the Proprietary license.

Where can I use Qwen3 Max?

Qwen3 Max is available through 2 providers including Novita, DeepInfra.

Where is the Qwen3 Max paper or technical report?

Qwen3 Max has a paper or technical report available at https://arxiv.org/abs/2505.09388. Use that source for architecture, training, release and evaluation details.

What models should I compare Qwen3 Max against?

Common Qwen3 Max comparisons include Qwen3 Max vs Claude 3.5 Sonnet, Qwen3 Max vs LongCat-Flash-Thinking-2601, Qwen3 Max vs Nova 2 Pro. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.