The AI arena is free today

Open Superagent
DeepSeekReleased on Mar 25, 2025

DeepSeek-V3 0324: Benchmarks, Pricing & Context Window

DeepSeek-V3 0324 is a language model from DeepSeek, released in March 2025, with a 164K-token context window, and pricing from $0.240/M input, $0.135/M cached input, $0.900/M output.

A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token). Features Multi-head Latent Attention (MLA), auxiliary-loss-free load balancing, and multi-token prediction training. Pre-trained on

Input
Text
Output
Text

DeepSeek-V3 0324 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for DeepSeek-V3 0324 across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How DeepSeek-V3 0324 holds up as conversations get longer.

Quality Tracker

DeepSeek-V3 0324 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Thu Oct 08 2026
Notice missing or incorrect data?

DeepSeek-V3 0324 pricing

Providers

DeepSeek-V3 0324 starts at $0.240 per million input tokens and $0.900 per million output tokens via DeepInfra. Reused prompt prefixes cost $0.135 per million cached input tokens. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.240$0.135$0.900163.8K/163.8K
—
—
/
Novita logoNovita
$0.280—$1.14163.8K/163.8K
—
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

DeepSeek-V3 0324 model size

DeepSeek-V3 0324 has 671 billion parameters and was trained on 14.8 trillion tokens. See how it compares to other models in the same parameter range.

ParametersTraining tokens
671B
14.8Ttokens
22× tokens-to-params ratio
Frontier (200B+)
671B
1B7B70B405B

DeepSeek-V3 0324 context window

Input and output token limits for DeepSeek-V3 0324, plus how it ranks on long-context understanding.

InputOutput
164Ktokens
164Ktokens
≈ 246 pages of text
164K
8K128K1M

Try now

huggle
DeepSeek-V3 0324in Huggle

Make it with
DeepSeek-V3 0324.

DeepSeek-V3 0324

DeepSeek-V3 0324 latency

DeepSeek-V3 0324 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

DeepSeek-V3 0324 examples

Recent arena outputs from DeepSeek-V3 0324, picked from the highest-ranked matchups.

DeepSeek-V3 0324 license

DeepSeek-V3 0324 is released under the MIT + Model License (Commercial use allowed) license, which restricts commercial use, has 671.0B parameters.

License
MIT + Model License (Commercial use allowed)
Non-commercial
Parameters
671.0B

DeepSeek-V3 0324 resources

Official sources for DeepSeek-V3 0324: provider documentation, official playground, paper or system card, source repository, model weights.

DeepSeek-V3 0324 vs other models

The most-compared alternatives to DeepSeek-V3 0324 are Phi 4 Reasoning Plus, DeepSeek R1 Zero, Nova 2 Pro. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like DeepSeek-V3 0324

Models ranked just above and below DeepSeek-V3 0324 by LLM Stats score.

 

Phi 4 Reasoning Plus

Score pending
 

DeepSeek R1 Zero

Score pending
 

Nova 2 Pro

Score pending
 

LongCat-Flash-Chat

Score pending
 

Claude 3.5 Sonnet

Score pending
 

Qwen3 VL 32B Instruct

Score pending

FAQ

Common questions about DeepSeek-V3 0324.

When was DeepSeek-V3 0324 released?

DeepSeek-V3 0324 was released on March 25, 2025 by DeepSeek. This is the official DeepSeek-V3 0324 release date tracked on LLM Stats.

How much does DeepSeek-V3 0324 cost?

DeepSeek-V3 0324 pricing starts at $0.24 per million input tokens, $0.14 per million cached input tokens, $0.90 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is DeepSeek-V3 0324?

DeepSeek-V3 0324 has 671 billion parameters. It was trained on 14.8 trillion tokens. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created DeepSeek-V3 0324?

DeepSeek-V3 0324 was created by DeepSeek.

What is the license for DeepSeek-V3 0324?

DeepSeek-V3 0324 is released under the MIT + Model License (Commercial use allowed) license. This is an open-source / open-weight license that permits self-hosting.

Where can I use DeepSeek-V3 0324?

DeepSeek-V3 0324 is available through 2 providers including DeepInfra, Novita.

Where is the DeepSeek-V3 0324 paper or technical report?

DeepSeek-V3 0324 has a paper or technical report available at https://arxiv.org/abs/2412.19437. Use that source for architecture, training, release and evaluation details.

What models should I compare DeepSeek-V3 0324 against?

Common DeepSeek-V3 0324 comparisons include DeepSeek-V3 0324 vs Phi 4 Reasoning Plus, DeepSeek-V3 0324 vs DeepSeek R1 Zero, DeepSeek-V3 0324 vs Nova 2 Pro. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.