The AI arena is free today

Open Superagent
DeepSeekReleased on Jan 20, 2025

DeepSeek R1 Distill Qwen 32B: API Pricing, Context Window & Benchmarks

DeepSeek R1 Distill Qwen 32B is a language model from DeepSeek, released in January 2025.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning

Input
Text
Output
Text

DeepSeek R1 Distill Qwen 32B benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for DeepSeek R1 Distill Qwen 32B across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How DeepSeek R1 Distill Qwen 32B holds up as conversations get longer.

Quality Tracker

DeepSeek R1 Distill Qwen 32B Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Sat Aug 29 2026
Notice missing or incorrect data?

DeepSeek R1 Distill Qwen 32B pricing

Providers

DeepSeek R1 Distill Qwen 32B starts at $0.120 per million input tokens and $0.180 per million output tokens via DeepInfra.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.120$0.180128.0K/128.0K
0.65
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

DeepSeek R1 Distill Qwen 32B model size

DeepSeek R1 Distill Qwen 32B has 32.8 billion parameters and was trained on 14.8 trillion tokens. See how it compares to other models in the same parameter range.

ParametersTraining tokens
32.8B
14.8Ttokens
451× tokens-to-params ratio
Large (30–80B)
32.8B
1B7B70B405B

DeepSeek R1 Distill Qwen 32B context window

Input and output token limits for DeepSeek R1 Distill Qwen 32B, plus how it ranks on long-context understanding.

InputOutput
128Ktokens
128Ktokens
192 pages of text
128K
8K128K1M

DeepSeek R1 Distill Qwen 32B API

Available from the model provider

DeepSeek R1 Distill Qwen 32B has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

DeepSeek R1 Distill Qwen 32B latency

DeepSeek R1 Distill Qwen 32B time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

DeepSeek R1 Distill Qwen 32B examples

Recent arena outputs from DeepSeek R1 Distill Qwen 32B, picked from the highest-ranked matchups.

DeepSeek R1 Distill Qwen 32B license

DeepSeek R1 Distill Qwen 32B is released under the MIT license, which permits commercial use, has 32.8B parameters.

License
MIT
Commercial use allowed
Parameters
32.8B

MIT License - allows commercial use

DeepSeek R1 Distill Qwen 32B resources

Official sources for DeepSeek R1 Distill Qwen 32B: api documentation, official playground, paper or system card, source repository, model weights.

DeepSeek R1 Distill Qwen 32B vs other models

The most-compared alternatives to DeepSeek R1 Distill Qwen 32B are Phi 4 Reasoning Plus, Nemotron Nano 9B v2, DeepSeek R1 Distill Llama 70B. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like DeepSeek R1 Distill Qwen 32B

Models ranked just above and below DeepSeek R1 Distill Qwen 32B by LLM Stats score.

 

Phi 4 Reasoning Plus

Score pending
 

Nemotron Nano 9B v2

Score pending
 

DeepSeek R1 Distill Llama 70B

Score pending
 

Ministral 3 (8B Reasoning 2512)

Score pending
 

o1-mini

Score pending
 

DeepSeek R1 Distill Qwen 14B

Score pending

FAQ

Common questions about DeepSeek R1 Distill Qwen 32B.

When was DeepSeek R1 Distill Qwen 32B released?

DeepSeek R1 Distill Qwen 32B was released on January 20, 2025 by DeepSeek. This is the official DeepSeek R1 Distill Qwen 32B release date tracked on LLM Stats.

How much does DeepSeek R1 Distill Qwen 32B cost?

DeepSeek R1 Distill Qwen 32B pricing starts at $0.12 per million input tokens and $0.18 per million output tokens via DeepInfra, the lowest price among tracked providers.

Is DeepSeek R1 Distill Qwen 32B available via API?

Yes, DeepSeek R1 Distill Qwen 32B is available via API. See the official documentation for authentication and endpoint details. It is served by 1 provider tracked on LLM Stats.

How big is DeepSeek R1 Distill Qwen 32B?

DeepSeek R1 Distill Qwen 32B has 32.8 billion parameters. It was trained on 14.8 trillion tokens. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created DeepSeek R1 Distill Qwen 32B?

DeepSeek R1 Distill Qwen 32B was created by DeepSeek.

What is the license for DeepSeek R1 Distill Qwen 32B?

DeepSeek R1 Distill Qwen 32B is released under the MIT license. This is an open-source / open-weight license that permits self-hosting.

What is DeepSeek R1 Distill Qwen 32B latency?

DeepSeek R1 Distill Qwen 32B p95 time to first token is 0.65 seconds via DeepInfra over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use DeepSeek R1 Distill Qwen 32B?

DeepSeek R1 Distill Qwen 32B is available through 1 provider including DeepInfra.

Where is the DeepSeek R1 Distill Qwen 32B paper or technical report?

DeepSeek R1 Distill Qwen 32B has a paper or technical report available at https://arxiv.org/pdf/2501.12948. Use that source for architecture, training, release and evaluation details.

What models should I compare DeepSeek R1 Distill Qwen 32B against?

Common DeepSeek R1 Distill Qwen 32B comparisons include DeepSeek R1 Distill Qwen 32B vs Phi 4 Reasoning Plus, DeepSeek R1 Distill Qwen 32B vs Nemotron Nano 9B v2, DeepSeek R1 Distill Qwen 32B vs DeepSeek R1 Distill Llama 70B. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.