The AI arena is free today

Open Superagent
DeepSeekReleased on Aug 21, 2026

DeepSeek-V4-Flash-Vision-Exp: Benchmarks, Pricing & Context Window

DeepSeek-V4-Flash-Vision-Exp is a language model from DeepSeek, released in August 2026, with multimodal input, a 1.0M-token context window, and pricing from $0.220/M input, $0.007/M cached input, $0.660/M output.

DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal vision-understanding model on the DeepSeek API (`model='deepseek-v4-flash-vision-exp'`). It accepts text and image inputs and produces text. DeepSeek reports pure-text agent,

Input
TextImage
Output
Text

DeepSeek-V4-Flash-Vision-Exp benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for DeepSeek-V4-Flash-Vision-Exp across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How DeepSeek-V4-Flash-Vision-Exp holds up as conversations get longer.

Quality Tracker

DeepSeek-V4-Flash-Vision-Exp Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Wed Sep 09 2026
Notice missing or incorrect data?

DeepSeek-V4-Flash-Vision-Exp pricing

Providers

DeepSeek-V4-Flash-Vision-Exp starts at $0.220 per million input tokens and $0.660 per million output tokens via DeepSeek. Reused prompt prefixes cost $0.0070 per million cached input tokens. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepSeek logoDeepSeek
$0.220$0.0070$0.6601.0M/393.2K
/
DeepInfra logoDeepInfra
$0.440$0.140$1.321.0M/1.0M
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

DeepSeek-V4-Flash-Vision-Exp context window

Input and output token limits for DeepSeek-V4-Flash-Vision-Exp, plus how it ranks on long-context understanding.

InputOutput
1.0Mtokens
1.0Mtokens
1.6k pages of text
1.0M
8K128K1M

Try now

huggle
DeepSeek-V4-Flash-Vision-Expin Huggle

Make it with
DeepSeek-V4-Flash-Vision-Exp.

DeepSeek-V4-Flash-Vision-Exp

DeepSeek-V4-Flash-Vision-Exp latency

DeepSeek-V4-Flash-Vision-Exp time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

DeepSeek-V4-Flash-Vision-Exp examples

Recent arena outputs from DeepSeek-V4-Flash-Vision-Exp, picked from the highest-ranked matchups.

DeepSeek-V4-Flash-Vision-Exp resources

Official sources for DeepSeek-V4-Flash-Vision-Exp: provider documentation, official launch post.

DeepSeek-V4-Flash-Vision-Exp vs other models

The most-compared alternatives to DeepSeek-V4-Flash-Vision-Exp are Qwen3.8 Max, Muse Spark 1.2, GLM-5.3-Flash. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like DeepSeek-V4-Flash-Vision-Exp

Models ranked just above and below DeepSeek-V4-Flash-Vision-Exp by LLM Stats score.

 

Qwen3.8 Max

Score pending
 

Muse Spark 1.2

Score pending
 

GLM-5.3-Flash

Score pending
 

GPT-5.6 Sol

Score pending
 

Seed 2.1 Pro

Score pending
 

Claude Fable 5

Score pending

FAQ

Common questions about DeepSeek-V4-Flash-Vision-Exp.

When was DeepSeek-V4-Flash-Vision-Exp released?

DeepSeek-V4-Flash-Vision-Exp was released on August 21, 2026 by DeepSeek. This is the official DeepSeek-V4-Flash-Vision-Exp release date tracked on LLM Stats.

How much does DeepSeek-V4-Flash-Vision-Exp cost?

DeepSeek-V4-Flash-Vision-Exp pricing starts at $0.22 per million input tokens, $0.01 per million cached input tokens, $0.66 per million output tokens via DeepSeek, the lowest price among tracked providers.

Who created DeepSeek-V4-Flash-Vision-Exp?

DeepSeek-V4-Flash-Vision-Exp was created by DeepSeek.

Is DeepSeek-V4-Flash-Vision-Exp multimodal?

Yes, DeepSeek-V4-Flash-Vision-Exp is multimodal and can accept both text and images as input.

Where can I use DeepSeek-V4-Flash-Vision-Exp?

DeepSeek-V4-Flash-Vision-Exp is available through 2 providers including DeepSeek, DeepInfra.

Where is the DeepSeek-V4-Flash-Vision-Exp paper or technical report?

DeepSeek-V4-Flash-Vision-Exp has a paper or technical report available at https://api-docs.deepseek.com/updates/. Use that source for architecture, training, release and evaluation details.

What models should I compare DeepSeek-V4-Flash-Vision-Exp against?

Common DeepSeek-V4-Flash-Vision-Exp comparisons include DeepSeek-V4-Flash-Vision-Exp vs Qwen3.8 Max, DeepSeek-V4-Flash-Vision-Exp vs Muse Spark 1.2, DeepSeek-V4-Flash-Vision-Exp vs GLM-5.3-Flash. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.