The AI arena is free today

Open Superagent

Qwen2.5 VL 32B Instruct vs QwQ-32B-Preview

Qwen2.5 VL 32B Instruct and QwQ-32B-Preview are closely matched at 9.7 and 9.0 on the LLM Stats Score.

Alibaba Cloud / Qwen Team · Alibaba Cloud / Qwen Team · Updated for 2026

Which is better?

Qwen2.5 VL 32B Instruct and QwQ-32B-Preview are closely matched on the overall LLM Stats Score at 9.7 and 9.0.

In the 1 individual benchmarks reported for both models, QwQ-32B-Preview wins 1; this is a narrower head-to-head signal than the composite indexes.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Qwen2.5 VL 32B Instruct

  • you want the most recent training data — it shipped Feb 2025

Choose QwQ-32B-Preview

  • you value its reported benchmark strengths — it wins 1 of 1 exact shared results

At a glance

The differences that matter most.

Core performance indexes
9.7
#274
9.0
#278
8.4
#279
9.3
#268
12.4
#154
4.4
#213
Cost, coverage & limits
Benchmark wins
0 of 1
1 of 1
Input price
— / M
$0.15 / M
Output price
— / M
$0.20 / M
Context window
32,768

Capability indexes

Additional strengths measured across groups of related public benchmarks

1 shared
Index
Qwen2.5 VL 32B Instruct
QwQ-32B-Preview
11.5#240
8.9#256
Conservative TrueSkill rating · higher is betterHow scores work

Individual benchmarks

28 reported for Qwen2.5 VL 32B Instruct · 4 for QwQ-32B-Preview

1 shared

Qwen2.5 VL 32B Instruct outperforms in 0 benchmarks, while QwQ-32B-Preview is better at 1 benchmark (GPQA).

QwQ-32B-Preview significantly outperforms across most benchmarks.

Tue Sep 22 2026 • llm-stats.com

Human preference

Blind head-to-head votes and playground preference scores

Model Size

Parameter count comparison

1.0B diff

Qwen2.5 VL 32B Instruct has 1.0B more parameters than QwQ-32B-Preview, making it 3.1% larger.

Alibaba Cloud / Qwen Team
Qwen2.5 VL 32B Instruct
33.5Bparameters
Alibaba Cloud / Qwen Team
QwQ-32B-Preview
32.5Bparameters
33.5B
Qwen2.5 VL 32B Instruct
32.5B
QwQ-32B-Preview

Context Window

Maximum input and output token capacity

Only QwQ-32B-Preview specifies input context (32,768 tokens). Only QwQ-32B-Preview specifies output context (32,768 tokens).

Alibaba Cloud / Qwen Team
Qwen2.5 VL 32B Instruct
Input- tokens
Output- tokens
Alibaba Cloud / Qwen Team
QwQ-32B-Preview
Input32,768 tokens
Output32,768 tokens
Tue Sep 22 2026 • llm-stats.com

Input capabilities

Documented input modalities across available providers

Qwen2.5 VL 32B Instruct supports multimodal inputs, whereas QwQ-32B-Preview does not.

Qwen2.5 VL 32B Instruct can handle both text and other forms of data like images, making it suitable for multimodal applications.

Qwen2.5 VL 32B Instruct

Text
Images
Audio
Video

QwQ-32B-Preview

Text
Images
Audio
Video

License

Usage and distribution terms

Both models are licensed under Apache 2.0.

Both models share the same licensing terms, providing consistent usage rights.

Qwen2.5 VL 32B Instruct

Apache 2.0

Open weights

QwQ-32B-Preview

Apache 2.0

Open weights

Release Timeline

When each model was launched

Qwen2.5 VL 32B Instruct was released on 2025-02-28, while QwQ-32B-Preview was released on 2024-11-28.

Qwen2.5 VL 32B Instruct is 3 months newer than QwQ-32B-Preview.

Qwen2.5 VL 32B Instruct

Feb 28, 2025

1.6 years ago

3mo newer
QwQ-32B-Preview

Nov 28, 2024

1.8 years ago

Knowledge Cutoff

When training data ends

QwQ-32B-Preview has a documented knowledge cutoff of 2024-11-28, while Qwen2.5 VL 32B Instruct's cutoff date is not specified.

We can confirm QwQ-32B-Preview's training data extends to 2024-11-28, but cannot make a direct comparison without Qwen2.5 VL 32B Instruct's cutoff date.

Qwen2.5 VL 32B Instruct

QwQ-32B-Preview

Nov 2024

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Qwen2.5 VL 32B Instruct and QwQ-32B-Preview side-by-side, then vote on the output you prefer.

Qwen2.5 VL 32B Instruct
✓ Preferred
QwQ-32B-Preview
Open in Playground

FAQ

Common questions about Qwen2.5 VL 32B Instruct vs QwQ-32B-Preview.

Which is better, Qwen2.5 VL 32B Instruct or QwQ-32B-Preview?

Qwen2.5 VL 32B Instruct and QwQ-32B-Preview are closely matched on the LLM Stats Score at 9.7 and 9.0. Qwen2.5 VL 32B Instruct is made by Alibaba Cloud / Qwen Team and QwQ-32B-Preview is made by Alibaba Cloud / Qwen Team. The best choice depends on your use case — compare their capability indexes, individual benchmarks, pricing, and limits above.

How does Qwen2.5 VL 32B Instruct compare to QwQ-32B-Preview in benchmarks?

Qwen2.5 VL 32B Instruct scores DocVQA: 94.8%, Android Control Low_EM: 93.3%, HumanEval: 91.5%, ScreenSpot: 88.5%, MBPP: 84.0%. QwQ-32B-Preview scores MATH-500: 90.6%, GPQA: 65.2%, AIME 2024: 50.0%, LiveCodeBench: 50.0%.

What are the context window sizes for Qwen2.5 VL 32B Instruct and QwQ-32B-Preview?

Qwen2.5 VL 32B Instruct supports an unknown number of tokens and QwQ-32B-Preview supports 33K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Qwen2.5 VL 32B Instruct and QwQ-32B-Preview?

Key differences include LLM Stats Score (9.7 vs 9.0), multimodal support (yes vs no). See the full comparison above for benchmark-by-benchmark results.