QwenReleased on Apr 29, 2025

Qwen3 235B A22B: API Pricing, Context Window & Benchmarks

Qwen3 235B A22B is a language model from Qwen, released in April 2025.

Qwen3 235B A22B is a large language model developed by Alibaba, featuring a Mixture-of-Experts (MoE) architecture with 235 billion total parameters and 22 billion activated parameters. It achieves competitive results in benchmark

Input
Text
Output
Text

Qwen3 235B A22B benchmarks

Rankings

Quality Tracker

Qwen3 235B A22B Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Jul 21 2026
Notice missing or incorrect data?

Qwen3 235B A22B pricing

Providers

Qwen3 235B A22B starts at $0.100 per million input tokens and $0.100 per million output tokens via Fireworks. See all 4 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MOutput $/MContext in / outTTFT p50 / p95 sOutput avg / p5 c/sSuccess 7dModalities in / out
Fireworks logoFireworks
$0.100$0.100128.0K/128.0K
/0.78
68/
/
DeepInfra logoDeepInfra
$0.200$0.600128.0K/128.0K
/1.23
22/
/
Novita logoNovita
$0.200$0.800128.0K/128.0K
/1.02
39/
/
Together logoTogether
$0.200$0.600128.0K/128.0K
/0.79
24/
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.

Loading chart...
Loading chart...
Loading chart...

Qwen3 235B A22B model size

Qwen3 235B A22B has 235 billion parameters and was trained on 36 trillion tokens. See how it compares to other models in the same parameter range.

ParametersTraining tokens
235B
36Ttokens
153× tokens-to-params ratio
Frontier (200B+)
235B
1B7B70B405B

Qwen3 235B A22B context window

Input and output token limits for Qwen3 235B A22B, plus how it ranks on long-context understanding.

InputOutput
128Ktokens
128Ktokens
192 pages of text
128K
8K128K1M

Qwen3 235B A22B API

Available from the model provider

Qwen3 235B A22B has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

Qwen3 235B A22B latency

Qwen3 235B A22B time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Qwen3 235B A22B examples

Recent arena outputs from Qwen3 235B A22B, picked from the highest-ranked matchups.

Qwen3 235B A22B license

Qwen3 235B A22B is released under the Apache 2.0 license, which permits commercial use, has 235.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
235.0B

Apache License 2.0 - allows commercial use

Qwen3 235B A22B resources

Official sources for Qwen3 235B A22B: api documentation, official playground, source repository, model weights.

Qwen3 235B A22B vs other models

The most-compared alternatives to Qwen3 235B A22B are Claude 3 Opus, MiMo-V2.5-Pro, GPT-4 Turbo. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3 235B A22B

Models ranked just above and below Qwen3 235B A22B by LLM Stats score.

 

Claude 3 Opus

Score pending
 

MiMo-V2.5-Pro

Score pending
 

GPT-4 Turbo

Score pending
 

GPT-4o

Score pending
 

Nova Pro

Score pending
 

Llama 3.2 90B Instruct

Score pending

FAQ

Common questions about Qwen3 235B A22B.

When was Qwen3 235B A22B released?

Qwen3 235B A22B was released on April 29, 2025 by Qwen. This is the official Qwen3 235B A22B release date tracked on LLM Stats.

How much does Qwen3 235B A22B cost?

Qwen3 235B A22B pricing starts at $0.10 per million input tokens and $0.10 per million output tokens via Fireworks, the lowest price among tracked providers.

Is Qwen3 235B A22B available via API?

Yes, Qwen3 235B A22B is available via API. See the official documentation for authentication and endpoint details. It is served by 4 providers tracked on LLM Stats.

How big is Qwen3 235B A22B?

Qwen3 235B A22B has 235 billion parameters. It was trained on 36.0 trillion tokens. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Qwen3 235B A22B?

Qwen3 235B A22B was created by Qwen.

What is the license for Qwen3 235B A22B?

Qwen3 235B A22B is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

What is Qwen3 235B A22B latency?

Qwen3 235B A22B p95 time to first token is 0.78 seconds via Fireworks over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use Qwen3 235B A22B?

Qwen3 235B A22B is available through 4 providers including Fireworks, DeepInfra, Novita, and 1 more.

What models should I compare Qwen3 235B A22B against?

Common Qwen3 235B A22B comparisons include Qwen3 235B A22B vs Claude 3 Opus, Qwen3 235B A22B vs MiMo-V2.5-Pro, Qwen3 235B A22B vs GPT-4 Turbo. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.