The AI arena is free today

Open Superagent
QwenReleased on Feb 13, 2026

Qwen3 Max Thinking: Benchmarks, Pricing & Context Window

Qwen3 Max Thinking is a language model from Qwen, released in February 2026, with a 256K-token context window, and pricing from $1.20/M input, $0.240/M cached input, $6.00/M output.

The latest flagship reasoning model in the Qwen3 family. Further enhanced by multiple innovations like adaptive tool-use and advanced test-time scaling techniques

Input
Text
Output
Text

Qwen3 Max Thinking benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Qwen3 Max Thinking across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Qwen3 Max Thinking holds up as conversations get longer.

Quality Tracker

Qwen3 Max Thinking Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Wed Sep 09 2026
Notice missing or incorrect data?

Qwen3 Max Thinking pricing

Providers

Qwen3 Max Thinking starts at $1.20 per million input tokens and $6.00 per million output tokens via DeepInfra. Reused prompt prefixes cost $0.240 per million cached input tokens.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$1.20$0.240$6.00256.0K/256.0K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Qwen3 Max Thinking context window

Input and output token limits for Qwen3 Max Thinking, plus how it ranks on long-context understanding.

InputOutput
256Ktokens
256Ktokens
385 pages of text
256K
8K128K1M

Try now

huggle
Qwen3 Max Thinkingin Huggle

Make it with
Qwen3 Max Thinking.

Qwen3 Max Thinking

Qwen3 Max Thinking latency

Qwen3 Max Thinking time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Qwen3 Max Thinking examples

Recent arena outputs from Qwen3 Max Thinking, picked from the highest-ranked matchups.

Qwen3 Max Thinking license

Qwen3 Max Thinking is a proprietary model available under its provider's product and API terms, has 1.0T parameters.

License
Proprietary
Hosted access
Parameters
1.0T

Proprietary license - usage restrictions apply

Qwen3 Max Thinking resources

Official sources for Qwen3 Max Thinking: provider documentation, official playground.

Qwen3 Max Thinking vs other models

The most-compared alternatives to Qwen3 Max Thinking are Gemini 3 Pro, GPT-5 High, Gemma 4 31B. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3 Max Thinking

Models ranked just above and below Qwen3 Max Thinking by LLM Stats score.

 

Gemini 3 Pro

Score pending
 

GPT-5 High

Score pending
 

Gemma 4 31B

Score pending
 

Claude 3.7 Sonnet

Score pending
 

Qwen3.7 Max

Score pending
 

GPT-5

Score pending

FAQ

Common questions about Qwen3 Max Thinking.

When was Qwen3 Max Thinking released?

Qwen3 Max Thinking was released on February 13, 2026 by Qwen. This is the official Qwen3 Max Thinking release date tracked on LLM Stats.

How much does Qwen3 Max Thinking cost?

Qwen3 Max Thinking pricing starts at $1.20 per million input tokens, $0.24 per million cached input tokens, $6.00 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is Qwen3 Max Thinking?

Qwen3 Max Thinking has 1000 billion parameters.

Who created Qwen3 Max Thinking?

Qwen3 Max Thinking was created by Qwen.

What is the license for Qwen3 Max Thinking?

Qwen3 Max Thinking is released under the Proprietary license.

Where can I use Qwen3 Max Thinking?

Qwen3 Max Thinking is available through 1 provider including DeepInfra.

What models should I compare Qwen3 Max Thinking against?

Common Qwen3 Max Thinking comparisons include Qwen3 Max Thinking vs Gemini 3 Pro, Qwen3 Max Thinking vs GPT-5 High, Qwen3 Max Thinking vs Gemma 4 31B. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.