- Organizations
- Qwen
- Qwen2.5 7B Instruct
Qwen2.5 7B Instruct: Benchmarks, Pricing & Context Window
Qwen2.5 7B Instruct is a language model from Qwen, released in September 2024.
Qwen2.5-7B-Instruct is an instruction-tuned 7B parameter language model that excels at following instructions, generating long texts (over 8K tokens), understanding structured data, and generating structured outputs like JSON. The model
Qwen2.5 7B Instruct benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Qwen2.5 7B Instruct across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Qwen2.5 7B Instruct holds up as conversations get longer.
Quality Tracker
Qwen2.5 7B Instruct Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Qwen2.5 7B Instruct pricing
Providers
Qwen2.5 7B Instruct starts at $0.300 per million input tokens and $0.300 per million output tokens via Together.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.300 | — | $0.300 | 131.1K/8.2K | 0.50 | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Qwen2.5 7B Instruct model size
Qwen2.5 7B Instruct has 7.6 billion parameters and was trained on 18 trillion tokens. See how it compares to other models in the same parameter range.
Qwen2.5 7B Instruct context window
Input and output token limits for Qwen2.5 7B Instruct, plus how it ranks on long-context understanding.
Try now
Make it with
Qwen2.5 7B Instruct.
Qwen2.5 7B Instruct latency
Qwen2.5 7B Instruct time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Qwen2.5 7B Instruct examples
Recent arena outputs from Qwen2.5 7B Instruct, picked from the highest-ranked matchups.
Qwen2.5 7B Instruct license
Qwen2.5 7B Instruct is released under the Apache 2.0 license, which permits commercial use, has 7.6B parameters.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 7.6B
Apache License 2.0 - allows commercial use
Qwen2.5 7B Instruct resources
Official sources for Qwen2.5 7B Instruct: provider documentation, paper or system card, official launch post, source repository, model weights.
Qwen2.5 7B Instruct vs other models
The most-compared alternatives to Qwen2.5 7B Instruct are Llama 3.1 405B Instruct, GPT-4, Grok-2. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Qwen2.5 7B Instruct
Models ranked just above and below Qwen2.5 7B Instruct by LLM Stats score.
FAQ
Common questions about Qwen2.5 7B Instruct.