The AI arena is free today

Open Superagent
API Provider4 active models1 organization

DeepSeek: API pricing, speed & models

DeepSeek hosts 4 active AI models, with input pricing from $0.14 per 1M tokens, with median throughput of 85 characters/sec, and P95 time to first token of 3.79s, with 100% success rate over 7 days. Compare DeepSeek's API speed, pricing, and reliability against other inference providers.

First-party only
Median throughput85 c/s
Median TTFT1.24s
P95 TTFT3.79s
Success rate (7d)100%
From$0.14 /M tok

Catalog

Type
Price
6 models
Model
At a glance

DeepSeekpricing, performance & catalog

The citable facts about DeepSeek's 4 models — sourced from provider APIs and refreshed continuously.

Lowest price
DeepSeek-V4-Flash-Max at $0.140 per 1M input tokens
Highest median throughput
DeepSeek-V4-Pro-0813 at 94 chars/s
Lowest median TTFT
DeepSeek-V3.2 (Non-thinking) at 1.22s
Largest context
DeepSeek-V4-Flash-Vision-Exp at 1.0M tokens
Catalog
4 active models from 1 organization

FAQ

Common questions about DeepSeek.

What is DeepSeek?

DeepSeek is an API provider that hosts large language models. Active models: 4; From (input): $0.14 / 1M tok; Median throughput: 85 c/s; Median TTFT: 1.24s; Success rate (7d): 100%.

How many models does DeepSeek offer?

DeepSeek currently serves 4 active models out of 11 historical offerings on LLM Stats.

What is DeepSeek's API pricing?

DeepSeek input pricing starts from $0.14 per 1M tokens, with the most expensive offering at $1.32 per 1M tokens. See the Pricing tab above for the full per-model breakdown.

How fast is DeepSeek?

DeepSeek delivers a median throughput of 85 characters per second and P95 time to first token of 3.79s across its model catalog. See the Models tab for per-model throughput and TTFT breakdowns.

Is DeepSeek reliable?

DeepSeek has a 100% success rate across 61 API calls in the last 7 days, with a 0% error rate.

Does DeepSeek support function calling?

Yes. 6 of 4 models on DeepSeek support function calling (tool use). The Capabilities tab lists which specific models accept tool definitions.

Does DeepSeek support JSON mode and structured output?

Yes. 6 of 4 models on DeepSeek support structured output (JSON mode / schema-constrained generation). The Capabilities tab shows which specific models accept response_format or json_schema parameters.

Does DeepSeek offer batch inference?

Yes. 1 of 4 models on DeepSeek support batch inference for cheaper, asynchronous workloads.

Does DeepSeek support multimodal models?

Yes. DeepSeek's catalog includes 1 vision-capable models. See the Models and Capabilities tabs for the full per-model breakdown.

Whose models does DeepSeek host?

DeepSeek hosts models from DeepSeek. See the Models tab for the full catalog grouped by creator.

How do I start using DeepSeek?

Sign up at https://deepseek.com/ to get an API key, then call DeepSeek's API directly from your application. Most clients work out of the box by pointing the OpenAI SDK at DeepSeek's base URL with your key. Use the Models and Pricing tabs above to pick the right model for your latency, cost, and context-window requirements.