- Organizations
- Qwen
- Qwen3 14B
Qwen3 14B: Benchmarks, Pricing & Context Window
Qwen3 14B is a language model from Qwen, released in April 2025, with a 41K-token context window, and pricing from $0.120/M input and $0.240/M output.
Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning,
Qwen3 14B benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Qwen3 14B across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Qwen3 14B holds up as conversations get longer.
Quality Tracker
Qwen3 14B Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Qwen3 14B pricing
Providers
Qwen3 14B starts at $0.120 per million input tokens and $0.240 per million output tokens via DeepInfra.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.120 | — | $0.240 | 41.0K/41.0K | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Qwen3 14B model size
Qwen3 14B has 14 billion parameters. See how it compares to other models in the same parameter range.
Qwen3 14B context window
Input and output token limits for Qwen3 14B, plus how it ranks on long-context understanding.
Try now
Make it with
Qwen3 14B.
Qwen3 14B latency
Qwen3 14B time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Qwen3 14B examples
Recent arena outputs from Qwen3 14B, picked from the highest-ranked matchups.
Qwen3 14B license
Qwen3 14B is released under the Apache 2.0 license, which permits commercial use, has 14.0B parameters.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 14.0B
Apache License 2.0 - allows commercial use
Qwen3 14B resources
Official sources for Qwen3 14B: provider documentation, official playground, source repository, model weights.
Qwen3 14B vs other models
The most-compared alternatives to Qwen3 14B are Kimi-k1.5, Phi 4 Reasoning Plus, QwQ-32B. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Qwen3 14B
Models ranked just above and below Qwen3 14B by LLM Stats score.
FAQ
Common questions about Qwen3 14B.