The AI arena is free today

Open Superagent
QwenReleased on Apr 28, 2025

Qwen3 14B: Benchmarks, Pricing & Context Window

Qwen3 14B is a language model from Qwen, released in April 2025, with a 41K-token context window, and pricing from $0.120/M input and $0.240/M output.

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning,

Input
Text
Output
Text

Qwen3 14B benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Qwen3 14B across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Qwen3 14B holds up as conversations get longer.

Quality Tracker

Qwen3 14B Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Sep 08 2026
Notice missing or incorrect data?

Qwen3 14B pricing

Providers

Qwen3 14B starts at $0.120 per million input tokens and $0.240 per million output tokens via DeepInfra.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.120$0.24041.0K/41.0K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Qwen3 14B model size

Qwen3 14B has 14 billion parameters. See how it compares to other models in the same parameter range.

Parameters
14B
Medium (10–30B)
14B
1B7B70B405B

Qwen3 14B context window

Input and output token limits for Qwen3 14B, plus how it ranks on long-context understanding.

InputOutput
41Ktokens
41Ktokens
62 pages of text
41K
8K128K1M

Try now

huggle
Qwen3 14Bin Huggle

Make it with
Qwen3 14B.

Qwen3 14B

Qwen3 14B latency

Qwen3 14B time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Qwen3 14B examples

Recent arena outputs from Qwen3 14B, picked from the highest-ranked matchups.

Qwen3 14B license

Qwen3 14B is released under the Apache 2.0 license, which permits commercial use, has 14.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
14.0B

Apache License 2.0 - allows commercial use

Qwen3 14B resources

Official sources for Qwen3 14B: provider documentation, official playground, source repository, model weights.

Qwen3 14B vs other models

The most-compared alternatives to Qwen3 14B are Kimi-k1.5, Phi 4 Reasoning Plus, QwQ-32B. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Qwen3 14B

Models ranked just above and below Qwen3 14B by LLM Stats score.

 

Kimi-k1.5

Score pending
 

Phi 4 Reasoning Plus

Score pending
 

QwQ-32B

Score pending
 

Claude 3.7 Sonnet

Score pending
 

Qwen3 30B A3B

Score pending
 

Sarvam-105B

Score pending

FAQ

Common questions about Qwen3 14B.

When was Qwen3 14B released?

Qwen3 14B was released on April 28, 2025 by Qwen. This is the official Qwen3 14B release date tracked on LLM Stats.

How much does Qwen3 14B cost?

Qwen3 14B pricing starts at $0.12 per million input tokens and $0.24 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is Qwen3 14B?

Qwen3 14B has 14 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Qwen3 14B?

Qwen3 14B was created by Qwen.

What is the license for Qwen3 14B?

Qwen3 14B is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

Where can I use Qwen3 14B?

Qwen3 14B is available through 1 provider including DeepInfra.

What models should I compare Qwen3 14B against?

Common Qwen3 14B comparisons include Qwen3 14B vs Kimi-k1.5, Qwen3 14B vs Phi 4 Reasoning Plus, Qwen3 14B vs QwQ-32B. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.