- Organizations
- DeepSeek
- DeepSeek R1 Distill Qwen 32B
DeepSeek R1 Distill Qwen 32B: API Pricing, Context Window & Benchmarks
DeepSeek R1 Distill Qwen 32B is a language model from DeepSeek, released in January 2025.
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning
DeepSeek R1 Distill Qwen 32B benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for DeepSeek R1 Distill Qwen 32B across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How DeepSeek R1 Distill Qwen 32B holds up as conversations get longer.
Quality Tracker
DeepSeek R1 Distill Qwen 32B Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
DeepSeek R1 Distill Qwen 32B pricing
Providers
DeepSeek R1 Distill Qwen 32B starts at $0.120 per million input tokens and $0.180 per million output tokens via DeepInfra.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.120 | — | $0.180 | 128.0K/128.0K | 0.65 | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
DeepSeek R1 Distill Qwen 32B model size
DeepSeek R1 Distill Qwen 32B has 32.8 billion parameters and was trained on 14.8 trillion tokens. See how it compares to other models in the same parameter range.
DeepSeek R1 Distill Qwen 32B context window
Input and output token limits for DeepSeek R1 Distill Qwen 32B, plus how it ranks on long-context understanding.
DeepSeek R1 Distill Qwen 32B API
Available from the model provider
DeepSeek R1 Distill Qwen 32B has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationDeepSeek R1 Distill Qwen 32B latency
DeepSeek R1 Distill Qwen 32B time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
DeepSeek R1 Distill Qwen 32B examples
Recent arena outputs from DeepSeek R1 Distill Qwen 32B, picked from the highest-ranked matchups.
DeepSeek R1 Distill Qwen 32B license
DeepSeek R1 Distill Qwen 32B is released under the MIT license, which permits commercial use, has 32.8B parameters.
- License
- MIT
- Commercial use allowed
- Parameters
- 32.8B
MIT License - allows commercial use
DeepSeek R1 Distill Qwen 32B resources
Official sources for DeepSeek R1 Distill Qwen 32B: api documentation, official playground, paper or system card, source repository, model weights.
DeepSeek R1 Distill Qwen 32B vs other models
The most-compared alternatives to DeepSeek R1 Distill Qwen 32B are Phi 4 Reasoning Plus, Nemotron Nano 9B v2, DeepSeek R1 Distill Llama 70B. Open any pair side-by-side for benchmarks, pricing, context, and latency.
- DeepSeek R1 Distill Qwen 32BvsPhi 4 Reasoning Plus
- DeepSeek R1 Distill Qwen 32BvsNemotron Nano 9B v2
- DeepSeek R1 Distill Qwen 32BvsDeepSeek R1 Distill Llama 70B
- DeepSeek R1 Distill Qwen 32BvsMinistral 3 (8B Reasoning 2512)
- DeepSeek R1 Distill Qwen 32Bvso1-mini
- DeepSeek R1 Distill Qwen 32BvsDeepSeek R1 Distill Qwen 14B
Models like DeepSeek R1 Distill Qwen 32B
Models ranked just above and below DeepSeek R1 Distill Qwen 32B by LLM Stats score.
FAQ
Common questions about DeepSeek R1 Distill Qwen 32B.