Model Comparison
Gemma 4 31B vs Qwen3.5-35B-A3BWhich is better in 2026?
Gemma 4 31B shows notably better performance in the majority of benchmarks. Gemma 4 31B is 3.6x cheaper per token.
Verdict: Gemma 4 31B vs Qwen3.5-35B-A3B — which is better?
Gemma 4 31B (by Google) and Qwen3.5-35B-A3B (by Alibaba Cloud / Qwen Team) are two of the AI models people compare most. Here is how they stack up on benchmarks, price and capabilities, and which one to pick in 2026.
Gemma 4 31B outperforms in 6 benchmarks (GPQA, LiveCodeBench v6, MathVision, MMMLU, MMMU-Pro, t2-bench), while Qwen3.5-35B-A3B is better at 3 benchmarks (Humanity's Last Exam, MedXpertQA, MMLU-Pro). Gemma 4 31B shows notably better performance in the majority of benchmarks.
On price, Gemma 4 31B is roughly 3.6x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.
Choose Gemma 4 31B if…
- you want the strongest raw capability — it leads on 6 of 9 shared benchmarks
- cost matters — it's about 3.6x cheaper per token
- you want the most recent training data — it shipped Apr 2026
Choose Qwen3.5-35B-A3B if…
- you want predictable pricing at $0.25/M input and $2.00/M output
Performance Benchmarks
Comparative analysis across standard metrics
Gemma 4 31B outperforms in 6 benchmarks (GPQA, LiveCodeBench v6, MathVision, MMMLU, MMMU-Pro, t2-bench), while Qwen3.5-35B-A3B is better at 3 benchmarks (Humanity's Last Exam, MedXpertQA, MMLU-Pro).
Gemma 4 31B shows notably better performance in the majority of benchmarks.
Arena Performance
Human preference votes
Pricing Analysis
Price comparison per million tokens
For input processing, Gemma 4 31B ($0.13/1M tokens) is 1.9x cheaper than Qwen3.5-35B-A3B ($0.25/1M tokens).
For output processing, Gemma 4 31B ($0.38/1M tokens) is 5.3x cheaper than Qwen3.5-35B-A3B ($2.00/1M tokens).
In conclusion, Qwen3.5-35B-A3B is more expensive than Gemma 4 31B.*
* Using a 3:1 ratio of input to output tokens
Model Size
Parameter count comparison
Qwen3.5-35B-A3B has 4.3B more parameters than Gemma 4 31B, making it 14.0% larger.
Context Window
Maximum input and output token capacity
Both models have the same input context window of 262,144 tokens. Gemma 4 31B can generate longer responses up to 131,072 tokens, while Qwen3.5-35B-A3B is limited to 65,000 tokens.
Input Capabilities
Supported data types and modalities
Both Gemma 4 31B and Qwen3.5-35B-A3B support multimodal inputs.
They are both capable of processing various types of data, offering versatility in application.
Gemma 4 31B
Qwen3.5-35B-A3B
License
Usage and distribution terms
Both models are licensed under Apache 2.0.
Both models share the same licensing terms, providing consistent usage rights.
Apache 2.0
Open weights
Apache 2.0
Open weights
Release Timeline
When each model was launched
Gemma 4 31B was released on 2026-04-02, while Qwen3.5-35B-A3B was released on 2026-02-24.
Gemma 4 31B is 1 month newer than Qwen3.5-35B-A3B.
Apr 2, 2026
3 months ago
1mo newerFeb 24, 2026
5 months ago
Knowledge Cutoff
When training data ends
Gemma 4 31B has a documented knowledge cutoff of 2025-01-01, while Qwen3.5-35B-A3B's cutoff date is not specified.
We can confirm Gemma 4 31B's training data extends to 2025-01-01, but cannot make a direct comparison without Qwen3.5-35B-A3B's cutoff date.
Jan 2025
—
Provider Availability
Gemma 4 31B is available from DeepInfra, FriendliAI, Novita, Together. Qwen3.5-35B-A3B is available from Novita.
Gemma 4 31B
Qwen3.5-35B-A3B
Outputs Comparison
Key Takeaways
Gemma 4 31B
View detailsQwen3.5-35B-A3B
View detailsAlibaba Cloud / Qwen Team
Detailed Comparison
Interactive Arena
Judge for yourself.
Run your own prompts against Gemma 4 31B and Qwen3.5-35B-A3B side-by-side, then vote on the output you prefer.
| Feature |
|---|
FAQ
Common questions about Gemma 4 31B vs Qwen3.5-35B-A3B.