The AI arena is free today

Open Superagent

Gemini 3.8 Flash Cyber vs Kimi K2 0905

Gemini 3.8 Flash Cyber leads the LLM Stats Score 42.4 to 21.8.

Google · Moonshot AI · Updated for 2026

Which is better?

Gemini 3.8 Flash Cyber leads the overall LLM Stats Score 42.4 to 21.8, ranking #46 overall.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Gemini 3.8 Flash Cyber

  • overall performance matters — it scores 42.4 and ranks #46 on LLM Stats
  • you want the most recent training data — it shipped Sep 2026

Choose Kimi K2 0905

  • you want predictable pricing at $0.60/M input and $2.50/M output

At a glance

The differences that matter most.

Core performance indexes
42.4
#46
21.8
#173
35.6
#25
19.5
#98
Cost, coverage & limits
Benchmark wins
Input price
— / M
$0.60 / M
Output price
— / M
$2.50 / M
Context window
262,144

Individual benchmarks

4 reported for Gemini 3.8 Flash Cyber · 6 for Kimi K2 0905

No common benchmarks found

Gemini 3.8 Flash Cyber and Kimi K2 0905don't have any common benchmark datasets to compare. They may have been evaluated on different testing suites.

Human preference

Blind head-to-head votes and playground preference scores

Context Window

Maximum input and output token capacity

Only Kimi K2 0905 specifies input context (262,144 tokens). Only Kimi K2 0905 specifies output context (262,144 tokens).

Google
Gemini 3.8 Flash Cyber
Input- tokens
Output- tokens
Moonshot AI
Kimi K2 0905
Input262,144 tokens
Output262,144 tokens
Fri Sep 04 2026 • llm-stats.com

License

Usage and distribution terms

Both models are licensed under proprietary licenses.

Both models have usage restrictions defined by their respective organizations.

Gemini 3.8 Flash Cyber

Proprietary

Closed source

Kimi K2 0905

Proprietary

Closed source

Release Timeline

When each model was launched

Gemini 3.8 Flash Cyber was released on 2026-09-02, while Kimi K2 0905 was released on 2025-09-05.

Gemini 3.8 Flash Cyber is 12 months newer than Kimi K2 0905.

Gemini 3.8 Flash Cyber

Sep 2, 2026

2 days ago

12mo newer
Kimi K2 0905

Sep 5, 2025

12 months ago

Knowledge Cutoff

When training data ends

Neither model specifies a knowledge cutoff date.

Unable to compare the recency of their training data.

No cutoff dates available

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Gemini 3.8 Flash Cyber and Kimi K2 0905 side-by-side, then vote on the output you prefer.

Gemini 3.8 Flash Cyber
✓ Preferred
Kimi K2 0905
Open in Playground

FAQ

Common questions about Gemini 3.8 Flash Cyber vs Kimi K2 0905.

Which is better, Gemini 3.8 Flash Cyber or Kimi K2 0905?

Gemini 3.8 Flash Cyber leads the LLM Stats Score 42.4 to 21.8. Gemini 3.8 Flash Cyber is made by Google and Kimi K2 0905 is made by Moonshot AI. The best choice depends on your use case — compare their capability indexes, individual benchmarks, pricing, and limits above.

How does Gemini 3.8 Flash Cyber compare to Kimi K2 0905 in benchmarks?

Gemini 3.8 Flash Cyber scores Gray Swan IPI Benchmark: 94.0%, CyberGym: 86.2%, Google Real-world Vulnerability Discovery: 71.0%, CWE-Bench: 47.2%. Kimi K2 0905 scores HumanEval: 94.5%, MMLU: 90.2%, MATH: 89.1%, MMLU-Pro: 82.5%, GPQA: 75.8%.

What are the context window sizes for Gemini 3.8 Flash Cyber and Kimi K2 0905?

Gemini 3.8 Flash Cyber supports an unknown number of tokens and Kimi K2 0905 supports 262K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Gemini 3.8 Flash Cyber and Kimi K2 0905?

Key differences include LLM Stats Score (42.4 vs 21.8). See the full comparison above for benchmark-by-benchmark results.

Who makes Gemini 3.8 Flash Cyber and Kimi K2 0905?

Gemini 3.8 Flash Cyber is developed by Google and Kimi K2 0905 is developed by Moonshot AI.