- Organizations
- Qwen
- Qwen3 VL 235B A22B Thinking
Qwen3 VL 235B A22B Thinking: API Pricing, Context Window & Benchmarks
Qwen3 VL 235B A22B Thinking is a language model from Qwen, released in September 2025, with multimodal input.
Qwen3-VL-235B-A22B-Thinking is the most powerful vision-language model in the Qwen series, featuring 236B parameters with MoE architecture for reasoning-enhanced multimodal understanding. Key capabilities include: Visual Agent (operates
Qwen3 VL 235B A22B Thinking benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
How Qwen3 VL 235B A22B Thinking performs across real-world prompt categories.
Performance by conversation depth
How Qwen3 VL 235B A22B Thinking holds up as conversations get longer.
Quality Tracker
Qwen3 VL 235B A22B Thinking Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Qwen3 VL 235B A22B Thinking pricing
Providers
Qwen3 VL 235B A22B Thinking starts at $0.450 per million input tokens and $3.49 per million output tokens via DeepInfra. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.450 | — | $3.49 | 262.1K/262.1K | — | — | / | |
| $0.980 | — | $3.95 | 131.1K/32.8K | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Qwen3 VL 235B A22B Thinking model size
Qwen3 VL 235B A22B Thinking has 236 billion parameters. See how it compares to other models in the same parameter range.
Qwen3 VL 235B A22B Thinking context window
Input and output token limits for Qwen3 VL 235B A22B Thinking, plus how it ranks on long-context understanding.
Qwen3 VL 235B A22B Thinking API
Available from the model provider
Qwen3 VL 235B A22B Thinking has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationQwen3 VL 235B A22B Thinking latency
Qwen3 VL 235B A22B Thinking time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Qwen3 VL 235B A22B Thinking examples
Recent arena outputs from Qwen3 VL 235B A22B Thinking, picked from the highest-ranked matchups.
Qwen3 VL 235B A22B Thinking license
Qwen3 VL 235B A22B Thinking is released under the Apache 2.0 license, which permits commercial use, has 236.0B parameters.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 236.0B
Apache License 2.0 - allows commercial use
Qwen3 VL 235B A22B Thinking resources
Official sources for Qwen3 VL 235B A22B Thinking: api documentation, official playground, paper or system card, official launch post, source repository, model weights.
Qwen3 VL 235B A22B Thinking vs other models
The most-compared alternatives to Qwen3 VL 235B A22B Thinking are GPT-5 Medium, Claude 3.5 Sonnet, K-EXAONE-236B-A23B. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Qwen3 VL 235B A22B Thinking
Models ranked just above and below Qwen3 VL 235B A22B Thinking by LLM Stats score.
FAQ
Common questions about Qwen3 VL 235B A22B Thinking.