The AI arena is free today

Open Superagent

Gemma 4 26B-A4B vs Inkling-Small

Inkling-Small significantly outperforms across most benchmarks. Gemma 4 26B-A4B is 2.7x cheaper per token.

Google · Thinking Machines Lab · Updated for 2026

Which is better?

Gemma 4 26B-A4B outperforms in 0 benchmarks, while Inkling-Small is better at 4 benchmarks (AIME 2026, GPQA, Humanity's Last Exam, MMMU-Pro). Inkling-Small significantly outperforms across most benchmarks.

On price, Gemma 4 26B-A4B is roughly 2.7x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.

Gemma 4 26B-A4B also accepts a larger context window (262,144 input tokens), making it the stronger choice for long documents and large codebases.

Based on current benchmark, pricing, and model metadata for 2026.

Choose Gemma 4 26B-A4B

  • cost matters — it's about 2.7x cheaper per token
  • you process long inputs — it offers a 262,144 token context window

Choose Inkling-Small

  • you want the strongest raw capability — it leads on 4 of 4 shared benchmarks
  • you want the most recent training data — it shipped Jul 2026

At a glance

The differences that matter most.

Benchmark wins
0 of 4
4 of 4
Input price
$0.13 / M
$0.30 / M
Output price
$0.40 / M
$1.20 / M
Context window
262,144
256,000
Released
Apr 2026
Jul 2026
License
Apache 2.0
Apache 2.0

Performance Benchmarks

Comparative analysis across standard metrics

4 benchmarks

Gemma 4 26B-A4B outperforms in 0 benchmarks, while Inkling-Small is better at 4 benchmarks (AIME 2026, GPQA, Humanity's Last Exam, MMMU-Pro).

Inkling-Small significantly outperforms across most benchmarks.

Tue Aug 25 2026 • llm-stats.com

Arena Performance

Playground indexes and blind preference scores

Pricing Analysis

Price comparison per million tokens

Gemma 4 26B-A4B costs less

For input processing, Gemma 4 26B-A4B ($0.13/1M tokens) is 2.3x cheaper than Inkling-Small ($0.30/1M tokens).

For output processing, Gemma 4 26B-A4B ($0.40/1M tokens) is 3.0x cheaper than Inkling-Small ($1.20/1M tokens).

In conclusion, Inkling-Small is more expensive than Gemma 4 26B-A4B.*

* Using a 3:1 ratio of input to output tokens

Lowest available price from all providers
Tue Aug 25 2026 • llm-stats.com
Google
Gemma 4 26B-A4B
Input tokens$0.13
Output tokens$0.40
Best providerNovita
Thinking Machines Lab
Inkling-Small
Input tokens$0.30
Output tokens$1.20
Best providerUnknown Organization
Notice missing or incorrect data?Start an Issue

Model Size

Parameter count comparison

250.8B diff

Inkling-Small has 250.8B more parameters than Gemma 4 26B-A4B, making it 995.2% larger.

Google
Gemma 4 26B-A4B
25.2Bparameters
Thinking Machines Lab
Inkling-Small
276.0Bparameters
25.2B
Gemma 4 26B-A4B
276.0B
Inkling-Small

Context Window

Maximum input and output token capacity

Gemma 4 26B-A4B accepts 262,144 input tokens compared to Inkling-Small's 256,000 tokens. Inkling-Small can generate longer responses up to 256,000 tokens, while Gemma 4 26B-A4B is limited to 131,072 tokens.

Google
Gemma 4 26B-A4B
Input262,144 tokens
Output131,072 tokens
Thinking Machines Lab
Inkling-Small
Input256,000 tokens
Output256,000 tokens
Tue Aug 25 2026 • llm-stats.com

Input Capabilities

Supported data types and modalities

Both Gemma 4 26B-A4B and Inkling-Small support multimodal inputs.

They are both capable of processing various types of data, offering versatility in application.

Gemma 4 26B-A4B

Text
Images
Audio
Video

Inkling-Small

Text
Images
Audio
Video

License

Usage and distribution terms

Both models are licensed under Apache 2.0.

Both models share the same licensing terms, providing consistent usage rights.

Gemma 4 26B-A4B

Apache 2.0

Open weights

Inkling-Small

Apache 2.0

Open weights

Release Timeline

When each model was launched

Gemma 4 26B-A4B was released on 2026-04-02, while Inkling-Small was released on 2026-07-30.

Inkling-Small is 4 months newer than Gemma 4 26B-A4B.

Gemma 4 26B-A4B

Apr 2, 2026

4 months ago

Inkling-Small

Jul 30, 2026

3 weeks ago

3mo newer

Knowledge Cutoff

When training data ends

Gemma 4 26B-A4B has a documented knowledge cutoff of 2025-01-01, while Inkling-Small's cutoff date is not specified.

We can confirm Gemma 4 26B-A4B's training data extends to 2025-01-01, but cannot make a direct comparison without Inkling-Small's cutoff date.

Gemma 4 26B-A4B

Jan 2025

Inkling-Small

Provider Availability

Gemma 4 26B-A4B is available from Novita. Inkling-Small is available from Thinking Machines Lab.

Gemma 4 26B-A4B

novita logo
Novita
Input Price:Input: $0.13/1MOutput Price:Output: $0.40/1M

Inkling-Small

thinking-machines logo
Unknown Organization
Input Price:Input: $0.30/1MOutput Price:Output: $1.20/1M
* Prices shown are per million tokens

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Gemma 4 26B-A4B and Inkling-Small side-by-side, then vote on the output you prefer.

Gemma 4 26B-A4B
✓ Preferred
Inkling-Small
Open in Playground

FAQ

Common questions about Gemma 4 26B-A4B vs Inkling-Small.

Which is better, Gemma 4 26B-A4B or Inkling-Small?

Inkling-Small significantly outperforms across most benchmarks. Gemma 4 26B-A4B is made by Google and Inkling-Small is made by Thinking Machines Lab. The best choice depends on your use case — compare their benchmark scores, pricing, and capabilities above.

How does Gemma 4 26B-A4B compare to Inkling-Small in benchmarks?

Gemma 4 26B-A4B scores AIME 2026: 88.3%, MMMLU: 86.3%, t2-bench: 85.5%, MMLU-Pro: 82.6%, MathVision: 82.4%. Inkling-Small scores AIME 2026: 95.5%, VoiceBench Avg: 90.1%, GPQA: 89.5%, Global-MMLU-Lite: 86.7%, ARC-AGI: 84.0%.

Is Gemma 4 26B-A4B cheaper than Inkling-Small?

Gemma 4 26B-A4B is 2.3x cheaper for input tokens. Gemma 4 26B-A4B costs $0.13/M input and $0.40/M output via novita. Inkling-Small costs $0.30/M input and $1.20/M output via thinking-machines.

What are the context window sizes for Gemma 4 26B-A4B and Inkling-Small?

Gemma 4 26B-A4B supports 262K tokens and Inkling-Small supports 256K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Gemma 4 26B-A4B and Inkling-Small?

Key differences include context window (262K vs 256K), input pricing ($0.13 vs $0.30/M). See the full comparison above for benchmark-by-benchmark results.

Who makes Gemma 4 26B-A4B and Inkling-Small?

Gemma 4 26B-A4B is developed by Google and Inkling-Small is developed by Thinking Machines Lab.