The AI arena is free today

Open Superagent

Mistral Large 2 vs Qwen2.5 32B Instruct

Mistral Large 2 and Qwen2.5 32B Instruct are closely matched at 7.4 and 7.7 on the LLM Stats Score.

Mistral AI · Alibaba Cloud / Qwen Team · Updated for 2026

Which is better?

Mistral Large 2 and Qwen2.5 32B Instruct are closely matched on the overall LLM Stats Score at 7.4 and 7.7.

In the 3 individual benchmarks reported for both models, Mistral Large 2 wins 2; this is a narrower head-to-head signal than the composite indexes.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Mistral Large 2

  • you value its reported benchmark strengths — it wins 2 of 3 exact shared results

Choose Qwen2.5 32B Instruct

  • you want the most recent training data — it shipped Sep 2024

At a glance

The differences that matter most.

Core performance indexes
7.4
#294
7.7
#290
7.3
#287
7.2
#289
13.0
#144
9.4
#178
Cost, coverage & limits
Benchmark wins
2 of 3
1 of 3
Input price
$2.00 / M
— / M
Output price
$6.00 / M
— / M
Context window
128,000
—

Capability indexes

Additional strengths measured across groups of related public benchmarks

1 shared
Index
Mistral Large 2
Qwen2.5 32B Instruct
11.3#245
16.4#207
Conservative TrueSkill rating · higher is betterHow scores work

Individual benchmarks

5 reported for Mistral Large 2 · 18 for Qwen2.5 32B Instruct

3 shared

Mistral Large 2 outperforms in 2 benchmarks (HumanEval, MMLU), while Qwen2.5 32B Instruct is better at 1 benchmark (GSM8k).

Mistral Large 2 shows notably better performance in the majority of benchmarks.

Sun Sep 27 2026 • llm-stats.com

Human preference

Blind head-to-head votes and playground preference scores

Model Size

Parameter count comparison

90.5B diff

Mistral Large 2 has 90.5B more parameters than Qwen2.5 32B Instruct, making it 278.5% larger.

Mistral AI
Mistral Large 2
123.0Bparameters
Alibaba Cloud / Qwen Team
Qwen2.5 32B Instruct
32.5Bparameters
123.0B
Mistral Large 2
32.5B
Qwen2.5 32B Instruct

Context Window

Maximum input and output token capacity

Only Mistral Large 2 specifies input context (128,000 tokens). Only Mistral Large 2 specifies output context (128,000 tokens).

Mistral AI
Mistral Large 2
Input128,000 tokens
Output128,000 tokens
Alibaba Cloud / Qwen Team
Qwen2.5 32B Instruct
Input- tokens
Output- tokens
Sun Sep 27 2026 • llm-stats.com

License

Usage and distribution terms

Mistral Large 2 is licensed under Mistral Research License, while Qwen2.5 32B Instruct uses Apache 2.0.

License differences may affect how you can use these models in commercial or open-source projects.

Mistral Large 2

Mistral Research License

Open weights

Qwen2.5 32B Instruct

Apache 2.0

Open weights

Release Timeline

When each model was launched

Mistral Large 2 was released on 2024-07-24, while Qwen2.5 32B Instruct was released on 2024-09-19.

Qwen2.5 32B Instruct is 2 months newer than Mistral Large 2.

Mistral Large 2

Jul 24, 2024

2.2 years ago

Qwen2.5 32B Instruct

Sep 19, 2024

2.0 years ago

1mo newer

Knowledge Cutoff

When training data ends

Neither model specifies a knowledge cutoff date.

Unable to compare the recency of their training data.

No cutoff dates available

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion→

Judge for yourself.

Run your own prompts against Mistral Large 2 and Qwen2.5 32B Instruct side-by-side, then vote on the output you prefer.

Mistral Large 2
✓ Preferred
Qwen2.5 32B Instruct
Open in Playground

FAQ

Common questions about Mistral Large 2 vs Qwen2.5 32B Instruct.

Which is better, Mistral Large 2 or Qwen2.5 32B Instruct?

Mistral Large 2 and Qwen2.5 32B Instruct are closely matched on the LLM Stats Score at 7.4 and 7.7. Mistral Large 2 is made by Mistral AI and Qwen2.5 32B Instruct is made by Alibaba Cloud / Qwen Team. The best choice depends on your use case — compare their capability indexes, individual benchmarks, pricing, and limits above.

How does Mistral Large 2 compare to Qwen2.5 32B Instruct in benchmarks?

Mistral Large 2 scores GSM8k: 93.0%, HumanEval: 92.0%, MT-Bench: 86.3%, MMLU: 84.0%, MMLU French: 82.8%. Qwen2.5 32B Instruct scores GSM8k: 95.9%, HumanEval: 88.4%, HellaSwag: 85.2%, BBH: 84.5%, MBPP: 84.0%.

What are the context window sizes for Mistral Large 2 and Qwen2.5 32B Instruct?

Mistral Large 2 supports 128K tokens and Qwen2.5 32B Instruct supports an unknown number of tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Mistral Large 2 and Qwen2.5 32B Instruct?

Key differences include LLM Stats Score (7.4 vs 7.7), licensing (Mistral Research License vs Apache 2.0). See the full comparison above for benchmark-by-benchmark results.

Who makes Mistral Large 2 and Qwen2.5 32B Instruct?

Mistral Large 2 is developed by Mistral AI and Qwen2.5 32B Instruct is developed by Alibaba Cloud / Qwen Team.