Model Comparison

Gemma 4 31B vs Qwen3.5-35B-A3BWhich is better in 2026?

Gemma 4 31B shows notably better performance in the majority of benchmarks. Gemma 4 31B is 3.6x cheaper per token.

Verdict: Gemma 4 31B vs Qwen3.5-35B-A3B — which is better?

Gemma 4 31B (by Google) and Qwen3.5-35B-A3B (by Alibaba Cloud / Qwen Team) are two of the AI models people compare most. Here is how they stack up on benchmarks, price and capabilities, and which one to pick in 2026.

Gemma 4 31B outperforms in 6 benchmarks (GPQA, LiveCodeBench v6, MathVision, MMMLU, MMMU-Pro, t2-bench), while Qwen3.5-35B-A3B is better at 3 benchmarks (Humanity's Last Exam, MedXpertQA, MMLU-Pro). Gemma 4 31B shows notably better performance in the majority of benchmarks.

On price, Gemma 4 31B is roughly 3.6x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.

Choose Gemma 4 31B if…

  • you want the strongest raw capability — it leads on 6 of 9 shared benchmarks
  • cost matters — it's about 3.6x cheaper per token
  • you want the most recent training data — it shipped Apr 2026

Choose Qwen3.5-35B-A3B if…

  • you want predictable pricing at $0.25/M input and $2.00/M output

Performance Benchmarks

Comparative analysis across standard metrics

9 benchmarks

Gemma 4 31B outperforms in 6 benchmarks (GPQA, LiveCodeBench v6, MathVision, MMMLU, MMMU-Pro, t2-bench), while Qwen3.5-35B-A3B is better at 3 benchmarks (Humanity's Last Exam, MedXpertQA, MMLU-Pro).

Gemma 4 31B shows notably better performance in the majority of benchmarks.

Tue Jul 28 2026 • llm-stats.com

Arena Performance

Human preference votes

Pricing Analysis

Price comparison per million tokens

Gemma 4 31B costs less

For input processing, Gemma 4 31B ($0.13/1M tokens) is 1.9x cheaper than Qwen3.5-35B-A3B ($0.25/1M tokens).

For output processing, Gemma 4 31B ($0.38/1M tokens) is 5.3x cheaper than Qwen3.5-35B-A3B ($2.00/1M tokens).

In conclusion, Qwen3.5-35B-A3B is more expensive than Gemma 4 31B.*

* Using a 3:1 ratio of input to output tokens

Lowest available price from all providers
Tue Jul 28 2026 • llm-stats.com
Google
Gemma 4 31B
Input tokens$0.13
Output tokens$0.38
Best providerDeepinfra
Alibaba Cloud / Qwen Team
Qwen3.5-35B-A3B
Input tokens$0.25
Output tokens$2.00
Best providerNovita
Notice missing or incorrect data?Start an Issue

Model Size

Parameter count comparison

4.3B diff

Qwen3.5-35B-A3B has 4.3B more parameters than Gemma 4 31B, making it 14.0% larger.

Google
Gemma 4 31B
30.7Bparameters
Alibaba Cloud / Qwen Team
Qwen3.5-35B-A3B
35.0Bparameters
30.7B
Gemma 4 31B
35.0B
Qwen3.5-35B-A3B

Context Window

Maximum input and output token capacity

Both models have the same input context window of 262,144 tokens. Gemma 4 31B can generate longer responses up to 131,072 tokens, while Qwen3.5-35B-A3B is limited to 65,000 tokens.

Google
Gemma 4 31B
Input262,144 tokens
Output131,072 tokens
Alibaba Cloud / Qwen Team
Qwen3.5-35B-A3B
Input262,144 tokens
Output65,000 tokens
Tue Jul 28 2026 • llm-stats.com

Input Capabilities

Supported data types and modalities

Both Gemma 4 31B and Qwen3.5-35B-A3B support multimodal inputs.

They are both capable of processing various types of data, offering versatility in application.

Gemma 4 31B

Text
Images
Audio
Video

Qwen3.5-35B-A3B

Text
Images
Audio
Video

License

Usage and distribution terms

Both models are licensed under Apache 2.0.

Both models share the same licensing terms, providing consistent usage rights.

Gemma 4 31B

Apache 2.0

Open weights

Qwen3.5-35B-A3B

Apache 2.0

Open weights

Release Timeline

When each model was launched

Gemma 4 31B was released on 2026-04-02, while Qwen3.5-35B-A3B was released on 2026-02-24.

Gemma 4 31B is 1 month newer than Qwen3.5-35B-A3B.

Gemma 4 31B

Apr 2, 2026

3 months ago

1mo newer
Qwen3.5-35B-A3B

Feb 24, 2026

5 months ago

Knowledge Cutoff

When training data ends

Gemma 4 31B has a documented knowledge cutoff of 2025-01-01, while Qwen3.5-35B-A3B's cutoff date is not specified.

We can confirm Gemma 4 31B's training data extends to 2025-01-01, but cannot make a direct comparison without Qwen3.5-35B-A3B's cutoff date.

Gemma 4 31B

Jan 2025

Qwen3.5-35B-A3B

Provider Availability

Gemma 4 31B is available from DeepInfra, FriendliAI, Novita, Together. Qwen3.5-35B-A3B is available from Novita.

Gemma 4 31B

deepinfra logo
Deepinfra
Input Price:Input: $0.13/1MOutput Price:Output: $0.38/1M
friendli logo
FriendliAI
Input Price:Input: $0.14/1MOutput Price:Output: $0.40/1M
novita logo
Novita
Input Price:Input: $0.14/1MOutput Price:Output: $0.40/1M
together logo
Together
Input Price:Input: $0.39/1MOutput Price:Output: $0.97/1M

Qwen3.5-35B-A3B

novita logo
Novita
Input Price:Input: $0.25/1MOutput Price:Output: $2.00/1M
* Prices shown are per million tokens

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Key Takeaways

Less expensive input tokens
Less expensive output tokens
Higher GPQA score (84.3% vs 84.2%)
Higher LiveCodeBench v6 score (80.0% vs 74.6%)
Higher MathVision score (85.6% vs 83.9%)
Higher MMMLU score (88.4% vs 85.2%)
Higher MMMU-Pro score (76.9% vs 75.1%)
Higher t2-bench score (86.4% vs 81.2%)
Alibaba Cloud / Qwen Team

Qwen3.5-35B-A3B

View details

Alibaba Cloud / Qwen Team

Higher Humanity's Last Exam score (47.4% vs 26.5%)
Higher MedXpertQA score (61.4% vs 61.3%)
Higher MMLU-Pro score (85.3% vs 85.2%)

Detailed Comparison

Interactive Arena

Judge for yourself.

Run your own prompts against Gemma 4 31B and Qwen3.5-35B-A3B side-by-side, then vote on the output you prefer.

Gemma 4 31B
✓ Preferred
Qwen3.5-35B-A3B
Open in Playground
AI Model Comparison Table
Feature
Google
Gemma 4 31B
Alibaba Cloud / Qwen Team
Qwen3.5-35B-A3B

FAQ

Common questions about Gemma 4 31B vs Qwen3.5-35B-A3B.

Which is better, Gemma 4 31B or Qwen3.5-35B-A3B?

Gemma 4 31B shows notably better performance in the majority of benchmarks. Gemma 4 31B is made by Google and Qwen3.5-35B-A3B is made by Alibaba Cloud / Qwen Team. The best choice depends on your use case — compare their benchmark scores, pricing, and capabilities above.

How does Gemma 4 31B compare to Qwen3.5-35B-A3B in benchmarks?

Gemma 4 31B scores AIME 2026: 89.2%, MMMLU: 88.4%, t2-bench: 86.4%, MathVision: 85.6%, MMLU-Pro: 85.2%. Qwen3.5-35B-A3B scores CountBench: 97.8%, VLMsAreBlind: 97.0%, MMLU-Redux: 93.3%, V*: 92.7%, AI2D: 92.6%.

Is Gemma 4 31B cheaper than Qwen3.5-35B-A3B?

Gemma 4 31B is 1.9x cheaper for input tokens. Gemma 4 31B costs $0.13/M input and $0.38/M output via deepinfra. Qwen3.5-35B-A3B costs $0.25/M input and $2.00/M output via novita.

What are the context window sizes for Gemma 4 31B and Qwen3.5-35B-A3B?

Gemma 4 31B supports 262K tokens and Qwen3.5-35B-A3B supports 262K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Gemma 4 31B and Qwen3.5-35B-A3B?

Key differences include input pricing ($0.13 vs $0.25/M). See the full comparison above for benchmark-by-benchmark results.

Who makes Gemma 4 31B and Qwen3.5-35B-A3B?

Gemma 4 31B is developed by Google and Qwen3.5-35B-A3B is developed by Alibaba Cloud / Qwen Team.