- Organizations
- Qwen
- Qwen3-Next-80B-A3B-Thinking
Qwen3-Next-80B-A3B-Thinking: Benchmarks, Pricing & Context Window
Qwen3-Next-80B-A3B-Thinking is a language model from Qwen, released in September 2025.
Qwen3-Next-80B-A3B-Thinking is the thinking variant of the Qwen3-Next series, featuring the same groundbreaking architecture as the instruct model. Leveraging GSPO, it addresses stability and efficiency challenges of hybrid attention +
Qwen3-Next-80B-A3B-Thinking benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Qwen3-Next-80B-A3B-Thinking across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Qwen3-Next-80B-A3B-Thinking holds up as conversations get longer.
Quality Tracker
Qwen3-Next-80B-A3B-Thinking Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Qwen3-Next-80B-A3B-Thinking pricing
Providers
Qwen3-Next-80B-A3B-Thinking starts at $0.150 per million input tokens and $1.50 per million output tokens via Novita.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.150 | — | $1.50 | 65.5K/65.5K | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Qwen3-Next-80B-A3B-Thinking model size
Qwen3-Next-80B-A3B-Thinking has 80 billion parameters and was trained on 15 trillion tokens. See how it compares to other models in the same parameter range.
Qwen3-Next-80B-A3B-Thinking context window
Input and output token limits for Qwen3-Next-80B-A3B-Thinking, plus how it ranks on long-context understanding.
Try now
Make it with
Qwen3-Next-80B-A3B-Thinking.
Qwen3-Next-80B-A3B-Thinking latency
Qwen3-Next-80B-A3B-Thinking time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Qwen3-Next-80B-A3B-Thinking examples
Recent arena outputs from Qwen3-Next-80B-A3B-Thinking, picked from the highest-ranked matchups.
Qwen3-Next-80B-A3B-Thinking license
Qwen3-Next-80B-A3B-Thinking is released under the Apache 2.0 license, which permits commercial use, has 80.0B parameters.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 80.0B
Apache License 2.0 - allows commercial use
Qwen3-Next-80B-A3B-Thinking resources
Official sources for Qwen3-Next-80B-A3B-Thinking: provider documentation, official playground, official launch post, source repository.
Qwen3-Next-80B-A3B-Thinking vs other models
The most-compared alternatives to Qwen3-Next-80B-A3B-Thinking are GPT-5 Medium, LongCat-Flash-Thinking, Llama 3.1 Nemotron Ultra 253B v1. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Qwen3-Next-80B-A3B-Thinking
Models ranked just above and below Qwen3-Next-80B-A3B-Thinking by LLM Stats score.
FAQ
Common questions about Qwen3-Next-80B-A3B-Thinking.