GLM-4.5-Air vs Kimi K2 Instruct
GLM-4.5-Air and Kimi K2 Instruct are closely matched at 24.6 and 22.0 on the LLM Stats Score.
Zhipu AI · Moonshot AI · Updated for 2026
Which is better?
GLM-4.5-Air and Kimi K2 Instruct are closely matched on the overall LLM Stats Score at 24.6 and 22.0.
In the 6 individual benchmarks reported for both models, GLM-4.5-Air wins 4; this is a narrower head-to-head signal than the composite indexes.
Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.
Choose GLM-4.5-Air
- you value its reported benchmark strengths — it wins 4 of 6 exact shared results
- you want the most recent training data — it shipped Jul 2025
Choose Kimi K2 Instruct
- you want predictable pricing at $0.50/M input and $0.50/M output
At a glance
The differences that matter most.
Capability indexes
Additional strengths measured across groups of related public benchmarks
Individual benchmarks
14 reported for GLM-4.5-Air · 38 for Kimi K2 Instruct
GLM-4.5-Air outperforms in 4 benchmarks (AIME 2024, Humanity's Last Exam, MATH-500, MMLU-Pro), while Kimi K2 Instruct is better at 1 benchmark (GPQA).
GLM-4.5-Air shows notably better performance in the majority of benchmarks.
Human preference
Blind head-to-head votes and playground preference scores
Model Size
Parameter count comparison
Kimi K2 Instruct has 894.0B more parameters than GLM-4.5-Air, making it 843.4% larger.
Context Window
Maximum input and output token capacity
Only Kimi K2 Instruct specifies input context (200,000 tokens). Only Kimi K2 Instruct specifies output context (200,000 tokens).
License
Usage and distribution terms
Both models are licensed under MIT.
Both models share the same licensing terms, providing consistent usage rights.
MIT
Open weights
MIT
Open weights
Release Timeline
When each model was launched
GLM-4.5-Air was released on 2025-07-28, while Kimi K2 Instruct was released on 2025-07-11.
GLM-4.5-Air is 1 month newer than Kimi K2 Instruct.
Jul 28, 2025
1.1 years ago
2w newerJul 11, 2025
1.1 years ago
Knowledge Cutoff
When training data ends
Neither model specifies a knowledge cutoff date.
Unable to compare the recency of their training data.
Outputs Comparison
Judge for yourself.
Run your own prompts against GLM-4.5-Air and Kimi K2 Instruct side-by-side, then vote on the output you prefer.
FAQ
Common questions about GLM-4.5-Air vs Kimi K2 Instruct.