The AI arena is free today

Open Superagent

Inkling-Small vs Kimi K2.6

Kimi K2.6 significantly outperforms across most benchmarks. Inkling-Small is 2.7x cheaper per token.

Thinking Machines Lab · Moonshot AI · Updated for 2026

Which is better?

Inkling-Small outperforms in 1 benchmarks (Toolathlon), while Kimi K2.6 is better at 8 benchmarks (AIME 2026, BrowseComp, CharXiv-R, GPQA, Humanity's Last Exam, MMMU-Pro, SciCode, SWE-Bench Pro). Kimi K2.6 significantly outperforms across most benchmarks.

On price, Inkling-Small is roughly 2.7x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.

Kimi K2.6 also accepts a larger context window (262,144 input tokens), making it the stronger choice for long documents and large codebases.

Based on current benchmark, pricing, and model metadata for 2026.

Choose Inkling-Small

  • cost matters — it's about 2.7x cheaper per token
  • you want the most recent training data — it shipped Jul 2026

Choose Kimi K2.6

  • you want the strongest raw capability — it leads on 9 of 10 shared benchmarks
  • you process long inputs — it offers a 262,144 token context window

At a glance

The differences that matter most.

Benchmark wins
1 of 10
9 of 10
Input price
$0.30 / M
$0.75 / M
Output price
$1.20 / M
$3.50 / M
Context window
256,000
262,144
Released
Jul 2026
Apr 2026
License
Apache 2.0
Modified MIT License

Performance Benchmarks

Comparative analysis across standard metrics

10 benchmarks

Inkling-Small outperforms in 1 benchmarks (Toolathlon), while Kimi K2.6 is better at 8 benchmarks (AIME 2026, BrowseComp, CharXiv-R, GPQA, Humanity's Last Exam, MMMU-Pro, SciCode, SWE-Bench Pro).

Kimi K2.6 significantly outperforms across most benchmarks.

Tue Aug 25 2026 • llm-stats.com

Arena Performance

Playground indexes and blind preference scores

Pricing Analysis

Price comparison per million tokens

Inkling-Small costs less

For input processing, Inkling-Small ($0.30/1M tokens) is 2.5x cheaper than Kimi K2.6 ($0.75/1M tokens).

For output processing, Inkling-Small ($1.20/1M tokens) is 2.9x cheaper than Kimi K2.6 ($3.50/1M tokens).

In conclusion, Kimi K2.6 is more expensive than Inkling-Small.*

* Using a 3:1 ratio of input to output tokens

Lowest available price from all providers
Tue Aug 25 2026 • llm-stats.com
Thinking Machines Lab
Inkling-Small
Input tokens$0.30
Output tokens$1.20
Best providerUnknown Organization
Moonshot AI
Kimi K2.6
Input tokens$0.75
Output tokens$3.50
Best providerDeepinfra
Notice missing or incorrect data?Start an Issue

Model Size

Parameter count comparison

724.0B diff

Kimi K2.6 has 724.0B more parameters than Inkling-Small, making it 262.3% larger.

Thinking Machines Lab
Inkling-Small
276.0Bparameters
Moonshot AI
Kimi K2.6
1.0Tparameters
276.0B
Inkling-Small
1000.0B
Kimi K2.6

Context Window

Maximum input and output token capacity

Kimi K2.6 accepts 262,144 input tokens compared to Inkling-Small's 256,000 tokens. Inkling-Small can generate longer responses up to 256,000 tokens, while Kimi K2.6 is limited to 131,072 tokens.

Thinking Machines Lab
Inkling-Small
Input256,000 tokens
Output256,000 tokens
Moonshot AI
Kimi K2.6
Input262,144 tokens
Output131,072 tokens
Tue Aug 25 2026 • llm-stats.com

Input Capabilities

Supported data types and modalities

Both Inkling-Small and Kimi K2.6 support multimodal inputs.

They are both capable of processing various types of data, offering versatility in application.

Inkling-Small

Text
Images
Audio
Video

Kimi K2.6

Text
Images
Audio
Video

License

Usage and distribution terms

Inkling-Small is licensed under Apache 2.0, while Kimi K2.6 uses Modified MIT License.

License differences may affect how you can use these models in commercial or open-source projects.

Inkling-Small

Apache 2.0

Open weights

Kimi K2.6

Modified MIT License

Open weights

Release Timeline

When each model was launched

Inkling-Small was released on 2026-07-30, while Kimi K2.6 was released on 2026-04-20.

Inkling-Small is 3 months newer than Kimi K2.6.

Inkling-Small

Jul 30, 2026

3 weeks ago

3mo newer
Kimi K2.6

Apr 20, 2026

4 months ago

Knowledge Cutoff

When training data ends

Neither model specifies a knowledge cutoff date.

Unable to compare the recency of their training data.

No cutoff dates available

Provider Availability

Inkling-Small is available from Thinking Machines Lab. Kimi K2.6 is available from DeepInfra, Fireworks, Moonshot AI, Novita, Together.

Inkling-Small

thinking-machines logo
Unknown Organization
Input Price:Input: $0.30/1MOutput Price:Output: $1.20/1M

Kimi K2.6

deepinfra logo
Deepinfra
Input Price:Input: $0.75/1MOutput Price:Output: $3.50/1M
fireworks logo
Fireworks
Input Price:Input: $0.95/1MOutput Price:Output: $4.00/1M
moonshot logo
Unknown Organization
Input Price:Input: $0.95/1MOutput Price:Output: $4.00/1M
novita logo
Novita
Input Price:Input: $0.95/1MOutput Price:Output: $4.00/1M
together logo
Together
Input Price:Input: $1.20/1MOutput Price:Output: $4.50/1M
* Prices shown are per million tokens

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Inkling-Small and Kimi K2.6 side-by-side, then vote on the output you prefer.

Inkling-Small
✓ Preferred
Kimi K2.6
Open in Playground

FAQ

Common questions about Inkling-Small vs Kimi K2.6.

Which is better, Inkling-Small or Kimi K2.6?

Kimi K2.6 significantly outperforms across most benchmarks. Inkling-Small is made by Thinking Machines Lab and Kimi K2.6 is made by Moonshot AI. The best choice depends on your use case — compare their benchmark scores, pricing, and capabilities above.

How does Inkling-Small compare to Kimi K2.6 in benchmarks?

Inkling-Small scores AIME 2026: 95.5%, VoiceBench Avg: 90.1%, GPQA: 89.5%, Global-MMLU-Lite: 86.7%, ARC-AGI: 84.0%. Kimi K2.6 scores V*: 96.9%, AIME 2026: 96.4%, MathVision: 93.2%, HMMT Feb 26: 92.7%, GPQA: 90.5%.

Is Inkling-Small cheaper than Kimi K2.6?

Inkling-Small is 2.5x cheaper for input tokens. Inkling-Small costs $0.30/M input and $1.20/M output via thinking-machines. Kimi K2.6 costs $0.75/M input and $3.50/M output via deepinfra.

What are the context window sizes for Inkling-Small and Kimi K2.6?

Inkling-Small supports 256K tokens and Kimi K2.6 supports 262K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Inkling-Small and Kimi K2.6?

Key differences include context window (256K vs 262K), input pricing ($0.30 vs $0.75/M), licensing (Apache 2.0 vs Modified MIT License). See the full comparison above for benchmark-by-benchmark results.

Who makes Inkling-Small and Kimi K2.6?

Inkling-Small is developed by Thinking Machines Lab and Kimi K2.6 is developed by Moonshot AI.