GLM-4.7-Flash vs Mercury 2
GLM-4.7-Flash and Mercury 2 are closely matched at 23.8 and 23.7 on the LLM Stats Score. GLM-4.7-Flash is 2.5x cheaper per token.
Zhipu AI · Inception · Updated for 2026
Which is better?
GLM-4.7-Flash and Mercury 2 are closely matched on the overall LLM Stats Score at 23.8 and 23.7.
In the 2 individual benchmarks reported for both models, GLM-4.7-Flash wins 2; this is a narrower head-to-head signal than the composite indexes.
On price, GLM-4.7-Flash is roughly 2.5x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.
Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.
Choose GLM-4.7-Flash
- you value its reported benchmark strengths — it wins 2 of 2 exact shared results
- cost matters — it's about 2.5x cheaper per token
- you need open weights you can self-host or fine-tune
Choose Mercury 2
- you want the most recent training data — it shipped Feb 2026
At a glance
The differences that matter most.
Capability indexes
Additional strengths measured across groups of related public benchmarks
Individual benchmarks
6 reported for GLM-4.7-Flash · 6 for Mercury 2
GLM-4.7-Flash outperforms in 2 benchmarks (AIME 2025, GPQA), while Mercury 2 is better at 0 benchmarks.
GLM-4.7-Flash significantly outperforms across most benchmarks.
Human preference
Blind head-to-head votes and playground preference scores
Pricing Analysis
Price comparison per million tokens
For input processing, GLM-4.7-Flash ($0.07/1M tokens) is 3.6x cheaper than Mercury 2 ($0.25/1M tokens).
For output processing, GLM-4.7-Flash ($0.40/1M tokens) is 1.9x cheaper than Mercury 2 ($0.75/1M tokens).
In conclusion, Mercury 2 is more expensive than GLM-4.7-Flash.*
* Using a 3:1 ratio of input to output tokens
Context Window
Maximum input and output token capacity
Both models have the same input context window of 128,000 tokens. GLM-4.7-Flash can generate longer responses up to 16,384 tokens, while Mercury 2 is limited to 8,192 tokens.
License
Usage and distribution terms
GLM-4.7-Flash is licensed under MIT, while Mercury 2 uses a proprietary license.
License differences may affect how you can use these models in commercial or open-source projects.
MIT
Open weights
Proprietary
Closed source
Release Timeline
When each model was launched
GLM-4.7-Flash was released on 2026-01-19, while Mercury 2 was released on 2026-02-24.
Mercury 2 is 1 month newer than GLM-4.7-Flash.
Jan 19, 2026
7 months ago
Feb 24, 2026
6 months ago
1mo newerKnowledge Cutoff
When training data ends
Neither model specifies a knowledge cutoff date.
Unable to compare the recency of their training data.
Provider Availability
GLM-4.7-Flash is available from ZAI. Mercury 2 is available from Inception.
GLM-4.7-Flash
Mercury 2
Outputs Comparison
Judge for yourself.
Run your own prompts against GLM-4.7-Flash and Mercury 2 side-by-side, then vote on the output you prefer.
FAQ
Common questions about GLM-4.7-Flash vs Mercury 2.