QwenReleased on Sep 19, 2024

Qwen2.5 32B Instruct: API Pricing, Context Window & Benchmarks

Qwen2.5 32B Instruct is a language model from Qwen, released in September 2024.

Qwen2.5-32B-Instruct is an instruction-tuned 32 billion parameter language model, part of the Qwen2.5 series. It is designed to follow instructions, generate long texts (over 8K tokens), understand structured data (e.g., tables), and

Qwen2.5 32B Instruct benchmarks

Rankings

Quality Tracker

Qwen2.5 32B Instruct Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Aug 03 2026
Notice missing or incorrect data?

Qwen2.5 32B Instruct model size

Qwen2.5 32B Instruct has 32.5 billion parameters and was trained on 18 trillion tokens. See how it compares to other models in the same parameter range.

ParametersTraining tokens
32.5B
18Ttokens
554× tokens-to-params ratio
Large (30–80B)
32.5B
1B7B70B405B

Qwen2.5 32B Instruct API

Available from the model provider

Qwen2.5 32B Instruct has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

Qwen2.5 32B Instruct latency

Qwen2.5 32B Instruct time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Qwen2.5 32B Instruct examples

Recent arena outputs from Qwen2.5 32B Instruct, picked from the highest-ranked matchups.

Qwen2.5 32B Instruct license

Qwen2.5 32B Instruct is released under the Apache 2.0 license, which permits commercial use, has 32.5B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
32.5B

Apache License 2.0 - allows commercial use

Qwen2.5 32B Instruct resources

Official sources for Qwen2.5 32B Instruct: api documentation, official launch post, source repository, model weights.

Qwen2.5 32B Instruct vs other models

The most-compared alternatives to Qwen2.5 32B Instruct are Claude 3 Opus, Llama 3.3 70B Instruct, Llama 3.1 70B Instruct. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen2.5 32B Instruct

Models ranked just above and below Qwen2.5 32B Instruct by LLM Stats score.

 

Claude 3 Opus

Score pending
 

Llama 3.3 70B Instruct

Score pending
 

Llama 3.1 70B Instruct

Score pending
 

Mistral Large 2

Score pending
 

GPT-5

Score pending
 

Mistral Small 3.2 24B Instruct

Score pending

FAQ

Common questions about Qwen2.5 32B Instruct.

When was Qwen2.5 32B Instruct released?

Qwen2.5 32B Instruct was released on September 19, 2024 by Qwen. This is the official Qwen2.5 32B Instruct release date tracked on LLM Stats.

Is Qwen2.5 32B Instruct available via API?

Yes, Qwen2.5 32B Instruct is available via API. See the official documentation for authentication and endpoint details.

How big is Qwen2.5 32B Instruct?

Qwen2.5 32B Instruct has 32.5 billion parameters. It was trained on 18.0 trillion tokens. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Qwen2.5 32B Instruct?

Qwen2.5 32B Instruct was created by Qwen.

What is the license for Qwen2.5 32B Instruct?

Qwen2.5 32B Instruct is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

Where is the Qwen2.5 32B Instruct paper or technical report?

Qwen2.5 32B Instruct has a paper or technical report available at https://qwenlm.github.io/blog/qwen2.5/. Use that source for architecture, training, release and evaluation details.

What models should I compare Qwen2.5 32B Instruct against?

Common Qwen2.5 32B Instruct comparisons include Qwen2.5 32B Instruct vs Claude 3 Opus, Qwen2.5 32B Instruct vs Llama 3.3 70B Instruct, Qwen2.5 32B Instruct vs Llama 3.1 70B Instruct. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.