Phi 4 Reasoning vs Qwen3 30B A3B
Phi 4 Reasoning and Qwen3 30B A3B are closely matched at 12.0 and 17.3 on the LLM Stats Score.
Microsoft · Alibaba Cloud / Qwen Team · Updated for 2026
Which is better?
Phi 4 Reasoning and Qwen3 30B A3B are closely matched on the overall LLM Stats Score at 12.0 and 17.3.
In the 5 individual benchmarks reported for both models, Qwen3 30B A3B wins 5; this is a narrower head-to-head signal than the composite indexes.
Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.
Choose Phi 4 Reasoning
- you want the most recent training data — it shipped Apr 2025
Choose Qwen3 30B A3B
- you value its reported benchmark strengths — it wins 5 of 5 exact shared results
At a glance
The differences that matter most.
Capability indexes
Additional strengths measured across groups of related public benchmarks
Individual benchmarks
11 reported for Phi 4 Reasoning · 8 for Qwen3 30B A3B
Phi 4 Reasoning outperforms in 0 benchmarks, while Qwen3 30B A3B is better at 4 benchmarks (AIME 2024, AIME 2025, Arena Hard, LiveCodeBench).
Qwen3 30B A3B significantly outperforms across most benchmarks.
Human preference
Blind head-to-head votes and playground preference scores
Model Size
Parameter count comparison
Qwen3 30B A3B has 16.5B more parameters than Phi 4 Reasoning, making it 117.9% larger.
Context Window
Maximum input and output token capacity
Only Qwen3 30B A3B specifies input context (128,000 tokens). Only Qwen3 30B A3B specifies output context (128,000 tokens).
License
Usage and distribution terms
Phi 4 Reasoning is licensed under MIT, while Qwen3 30B A3B uses Apache 2.0.
License differences may affect how you can use these models in commercial or open-source projects.
MIT
Open weights
Apache 2.0
Open weights
Release Timeline
When each model was launched
Phi 4 Reasoning was released on 2025-04-30, while Qwen3 30B A3B was released on 2025-04-29.
Phi 4 Reasoning is 0 month newer than Qwen3 30B A3B.
Apr 30, 2025
1.4 years ago
1d newerApr 29, 2025
1.4 years ago
Knowledge Cutoff
When training data ends
Phi 4 Reasoning has a documented knowledge cutoff of 2025-03-01, while Qwen3 30B A3B's cutoff date is not specified.
We can confirm Phi 4 Reasoning's training data extends to 2025-03-01, but cannot make a direct comparison without Qwen3 30B A3B's cutoff date.
Mar 2025
—
Outputs Comparison
Judge for yourself.
Run your own prompts against Phi 4 Reasoning and Qwen3 30B A3B side-by-side, then vote on the output you prefer.
FAQ
Common questions about Phi 4 Reasoning vs Qwen3 30B A3B.