DeepSeekReleased on Dec 1, 2025

DeepSeek-V3.2 (Thinking): API Pricing, Context Window & Benchmarks

DeepSeek-V3.2 (Thinking) is a language model from DeepSeek, released in December 2025.

DeepSeek-V3.2 in thinking mode. A powerful reasoning model with 685B parameters using DeepSeek Sparse Attention (DSA). This mode enables extended chain-of-thought reasoning for complex problem-solving tasks. Supports JSON output and tool

Input
Text
Output
Text

DeepSeek-V3.2 (Thinking) benchmarks

Rankings

Quality Tracker

DeepSeek-V3.2 (Thinking) Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Jul 20 2026
Notice missing or incorrect data?

DeepSeek-V3.2 (Thinking) pricing

Providers

DeepSeek-V3.2 (Thinking) starts at $0.280 per million input tokens and $0.420 per million output tokens via DeepSeek.

ProviderInput $/MOutput $/MWorkload 1M + 100KContext in / outTTFT p50 / p95 sOutput avg / p5 c/sSuccess 7dModalities in / out
DeepSeek logoDeepSeek
$0.280$0.420$0.322131.1K/65.5K
/0.50
50/
/

Workload cost uses 1M input tokens plus 100K output tokens at published list prices without assuming a cache hit. Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.

DeepSeek-V3.2 (Thinking) model size

DeepSeek-V3.2 (Thinking) has 685 billion parameters. See how it compares to other models in the same parameter range.

Parameters
685BMoE
Frontier (200B+)
685B
1B7B70B405B

DeepSeek-V3.2 (Thinking) context window

Input and output token limits for DeepSeek-V3.2 (Thinking), plus how it ranks on long-context understanding.

InputOutput
131Ktokens
66Ktokens
197 pages of text
131K
8K128K1M

DeepSeek-V3.2 (Thinking) API

Available from the model provider

DeepSeek-V3.2 (Thinking) has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

DeepSeek-V3.2 (Thinking) latency

DeepSeek-V3.2 (Thinking) time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

DeepSeek-V3.2 (Thinking) examples

Recent arena outputs from DeepSeek-V3.2 (Thinking), picked from the highest-ranked matchups.

DeepSeek-V3.2 (Thinking) license

DeepSeek-V3.2 (Thinking) is released under the MIT license, which permits commercial use, has 685.0B parameters.

License
MIT
Commercial use allowed
Parameters
685.0B

MIT License - allows commercial use

DeepSeek-V3.2 (Thinking) resources

Official sources for DeepSeek-V3.2 (Thinking): api documentation, official playground, paper or system card, model weights.

DeepSeek-V3.2 (Thinking) vs other models

The most-compared alternatives to DeepSeek-V3.2 (Thinking) are Grok-3, Seed 2.0 Lite, K-EXAONE-236B-A23B. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like DeepSeek-V3.2 (Thinking)

Models ranked just above and below DeepSeek-V3.2 (Thinking) by LLM Stats score.

 

Grok-3

Score pending
 

Seed 2.0 Lite

Score pending
 

K-EXAONE-236B-A23B

Score pending
 

Gemma 4 31B

Score pending
 

Gemma 4 26B-A4B

Score pending
 

Claude Sonnet 4

Score pending

FAQ

Common questions about DeepSeek-V3.2 (Thinking).

When was DeepSeek-V3.2 (Thinking) released?

DeepSeek-V3.2 (Thinking) was released on December 1, 2025 by DeepSeek. This is the official DeepSeek-V3.2 (Thinking) release date tracked on LLM Stats.

How much does DeepSeek-V3.2 (Thinking) cost?

DeepSeek-V3.2 (Thinking) pricing starts at $0.28 per million input tokens and $0.42 per million output tokens via DeepSeek, the lowest price among tracked providers.

Is DeepSeek-V3.2 (Thinking) available via API?

Yes, DeepSeek-V3.2 (Thinking) is available via API. See the official documentation for authentication and endpoint details. It is served by 1 provider tracked on LLM Stats.

How big is DeepSeek-V3.2 (Thinking)?

DeepSeek-V3.2 (Thinking) has 685 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created DeepSeek-V3.2 (Thinking)?

DeepSeek-V3.2 (Thinking) was created by DeepSeek.

What is the license for DeepSeek-V3.2 (Thinking)?

DeepSeek-V3.2 (Thinking) is released under the MIT license. This is an open-source / open-weight license that permits self-hosting.

What is DeepSeek-V3.2 (Thinking) latency?

DeepSeek-V3.2 (Thinking) p95 time to first token is 0.50 seconds via DeepSeek over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use DeepSeek-V3.2 (Thinking)?

DeepSeek-V3.2 (Thinking) is available through 1 provider including DeepSeek.

Where is the DeepSeek-V3.2 (Thinking) paper or technical report?

DeepSeek-V3.2 (Thinking) has a paper or technical report available at https://huggingface.co/deepseek-ai/DeepSeek-V3.2/resolve/main/assets/paper.pdf. Use that source for architecture, training, release and evaluation details.

What models should I compare DeepSeek-V3.2 (Thinking) against?

Common DeepSeek-V3.2 (Thinking) comparisons include DeepSeek-V3.2 (Thinking) vs Grok-3, DeepSeek-V3.2 (Thinking) vs Seed 2.0 Lite, DeepSeek-V3.2 (Thinking) vs K-EXAONE-236B-A23B. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.