The AI arena is free today

Open Superagent
QwenReleased on Jul 25, 2025

Qwen3-235B-A22B-Thinking-2507: Benchmarks, Pricing & Context Window

Qwen3-235B-A22B-Thinking-2507 is a language model from Qwen, released in July 2025.

Qwen3-235B-A22B-Thinking-2507 is a state-of-the-art thinking-enabled Mixture-of-Experts (MoE) model with 235B total parameters (22B activated). It features 94 layers, 128 experts (8 activated), and supports 262K native context length. This

Input
Text
Output
Text

Qwen3-235B-A22B-Thinking-2507 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Qwen3-235B-A22B-Thinking-2507 across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Qwen3-235B-A22B-Thinking-2507 holds up as conversations get longer.

Quality Tracker

Qwen3-235B-A22B-Thinking-2507 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Fri Oct 02 2026
Notice missing or incorrect data?

Qwen3-235B-A22B-Thinking-2507 pricing

Providers

Qwen3-235B-A22B-Thinking-2507 starts at $0.300 per million input tokens and $3.00 per million output tokens via Fireworks. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Fireworks logoFireworks
$0.300—$3.00262.1K/131.1K
—
—
/
Novita logoNovita
$0.300—$3.00256.0K/32.8K
—
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Qwen3-235B-A22B-Thinking-2507 model size

Qwen3-235B-A22B-Thinking-2507 has 235 billion parameters. See how it compares to other models in the same parameter range.

Parameters
235B
Frontier (200B+)
235B
1B7B70B405B

Qwen3-235B-A22B-Thinking-2507 context window

Input and output token limits for Qwen3-235B-A22B-Thinking-2507, plus how it ranks on long-context understanding.

InputOutput
262Ktokens
131Ktokens
≈ 394 pages of text
262K
8K128K1M

Try now

huggle
Qwen3-235B-A22B-Thinking-2507in Huggle

Make it with
Qwen3-235B-A22B-Thinking-2507.

Qwen3-235B-A22B-Thinking-2507

Qwen3-235B-A22B-Thinking-2507 latency

Qwen3-235B-A22B-Thinking-2507 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Qwen3-235B-A22B-Thinking-2507 examples

Recent arena outputs from Qwen3-235B-A22B-Thinking-2507, picked from the highest-ranked matchups.

Qwen3-235B-A22B-Thinking-2507 license

Qwen3-235B-A22B-Thinking-2507 is released under the Apache 2.0 license, which permits commercial use, has 235.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
235.0B

Apache License 2.0 - allows commercial use

Qwen3-235B-A22B-Thinking-2507 resources

Official sources for Qwen3-235B-A22B-Thinking-2507: provider documentation, official playground, official launch post, source repository.

Qwen3-235B-A22B-Thinking-2507 vs other models

The most-compared alternatives to Qwen3-235B-A22B-Thinking-2507 are K-EXAONE-236B-A23B, GPT OSS 120B High, Nova 2 Pro. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3-235B-A22B-Thinking-2507

Models ranked just above and below Qwen3-235B-A22B-Thinking-2507 by LLM Stats score.

 

K-EXAONE-236B-A23B

Score pending
 

GPT OSS 120B High

Score pending
 

Nova 2 Pro

Score pending
 

Grok 4 Fast

Score pending
 

Nova 2 Omni

Score pending
 

Gemini 2.5 Pro

Score pending

FAQ

Common questions about Qwen3-235B-A22B-Thinking-2507.

When was Qwen3-235B-A22B-Thinking-2507 released?

Qwen3-235B-A22B-Thinking-2507 was released on July 25, 2025 by Qwen. This is the official Qwen3-235B-A22B-Thinking-2507 release date tracked on LLM Stats.

How much does Qwen3-235B-A22B-Thinking-2507 cost?

Qwen3-235B-A22B-Thinking-2507 pricing starts at $0.30 per million input tokens and $3.00 per million output tokens via Fireworks, the lowest price among tracked providers.

How big is Qwen3-235B-A22B-Thinking-2507?

Qwen3-235B-A22B-Thinking-2507 has 235 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Qwen3-235B-A22B-Thinking-2507?

Qwen3-235B-A22B-Thinking-2507 was created by Qwen.

What is the license for Qwen3-235B-A22B-Thinking-2507?

Qwen3-235B-A22B-Thinking-2507 is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

Where can I use Qwen3-235B-A22B-Thinking-2507?

Qwen3-235B-A22B-Thinking-2507 is available through 2 providers including Fireworks, Novita.

Where is the Qwen3-235B-A22B-Thinking-2507 paper or technical report?

Qwen3-235B-A22B-Thinking-2507 has a paper or technical report available at https://qwenlm.github.io/blog/qwen3-thinking/. Use that source for architecture, training, release and evaluation details.

What models should I compare Qwen3-235B-A22B-Thinking-2507 against?

Common Qwen3-235B-A22B-Thinking-2507 comparisons include Qwen3-235B-A22B-Thinking-2507 vs K-EXAONE-236B-A23B, Qwen3-235B-A22B-Thinking-2507 vs GPT OSS 120B High, Qwen3-235B-A22B-Thinking-2507 vs Nova 2 Pro. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.