The AI arena is free today

Open Superagent
QwenReleased on Sep 22, 2025

Qwen3 VL 8B Instruct: API Pricing, Context Window & Benchmarks

Qwen3 VL 8B Instruct is a language model from Qwen, released in September 2025, with multimodal input.

Qwen3-VL is a large multimodal model that unifies vision, language, and reasoning to achieve human-level perception and cognition across text, images, and video. Built on a 235B-parameter architecture, it integrates early joint training of

Input
TextImageVideo
Output
Text

Qwen3 VL 8B Instruct benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

How Qwen3 VL 8B Instruct performs across real-world prompt categories.

Performance by conversation depth

How Qwen3 VL 8B Instruct holds up as conversations get longer.

Quality Tracker

Qwen3 VL 8B Instruct Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Aug 25 2026
Notice missing or incorrect data?

Qwen3 VL 8B Instruct pricing

Providers

Qwen3 VL 8B Instruct starts at $0.0800 per million input tokens and $0.500 per million output tokens via Novita. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Novita logoNovita
$0.0800$0.500131.1K/32.8K
/
DeepInfra logoDeepInfra
$0.180$0.690262.1K/262.1K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Qwen3 VL 8B Instruct model size

Qwen3 VL 8B Instruct has 9 billion parameters. See how it compares to other models in the same parameter range.

Parameters
9B
Small (3–10B)
9B
1B7B70B405B

Qwen3 VL 8B Instruct context window

Input and output token limits for Qwen3 VL 8B Instruct, plus how it ranks on long-context understanding.

InputOutput
262Ktokens
262Ktokens
394 pages of text
262K
8K128K1M

Qwen3 VL 8B Instruct API

Available from the model provider

Qwen3 VL 8B Instruct has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

Qwen3 VL 8B Instruct latency

Qwen3 VL 8B Instruct time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Qwen3 VL 8B Instruct examples

Recent arena outputs from Qwen3 VL 8B Instruct, picked from the highest-ranked matchups.

Qwen3 VL 8B Instruct license

Qwen3 VL 8B Instruct is released under the Apache 2.0 license, which permits commercial use, has 9.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
9.0B

Apache License 2.0 - allows commercial use

Qwen3 VL 8B Instruct resources

Official sources for Qwen3 VL 8B Instruct: api documentation, official playground, official launch post, source repository, model weights.

Qwen3 VL 8B Instruct vs other models

The most-compared alternatives to Qwen3 VL 8B Instruct are Grok-2 mini, Nova Lite, Phi 4. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3 VL 8B Instruct

Models ranked just above and below Qwen3 VL 8B Instruct by LLM Stats score.

 

Grok-2 mini

Score pending
 

Nova Lite

Score pending
 

Phi 4

Score pending
 

Llama 4 Maverick

Score pending
 

Qwen3 VL 4B Thinking

Score pending
 

Qwen2.5 72B Instruct

Score pending

FAQ

Common questions about Qwen3 VL 8B Instruct.

When was Qwen3 VL 8B Instruct released?

Qwen3 VL 8B Instruct was released on September 22, 2025 by Qwen. This is the official Qwen3 VL 8B Instruct release date tracked on LLM Stats.

How much does Qwen3 VL 8B Instruct cost?

Qwen3 VL 8B Instruct pricing starts at $0.08 per million input tokens and $0.50 per million output tokens via Novita, the lowest price among tracked providers.

Is Qwen3 VL 8B Instruct available via API?

Yes, Qwen3 VL 8B Instruct is available via API. See the official documentation for authentication and endpoint details. It is served by 2 providers tracked on LLM Stats.

How big is Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct has 9 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct was created by Qwen.

What is the license for Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

Is Qwen3 VL 8B Instruct multimodal?

Yes, Qwen3 VL 8B Instruct is multimodal and can accept both text and images as input.

Where can I use Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct is available through 2 providers including Novita, DeepInfra.

Where is the Qwen3 VL 8B Instruct paper or technical report?

Qwen3 VL 8B Instruct has a paper or technical report available at https://qwen.ai/blog?id=99f0335c4ad9ff6153e517418d48535ab6d8afef&from=research.latest-advancements-list. Use that source for architecture, training, release and evaluation details.

What models should I compare Qwen3 VL 8B Instruct against?

Common Qwen3 VL 8B Instruct comparisons include Qwen3 VL 8B Instruct vs Grok-2 mini, Qwen3 VL 8B Instruct vs Nova Lite, Qwen3 VL 8B Instruct vs Phi 4. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.