The AI arena is free today

Open Superagent
ZAIReleased on Dec 22, 2025

GLM-4.7: Benchmarks, Pricing & Context Window

GLM-4.7 is a language model from ZAI, released in December 2025, with multimodal input, a 203K-token context window, and pricing from $0.400/M input, $0.080/M cached input, $1.75/M output.

GLM 4.7 is a coding‑centric model that thinks before acting, preserves its reasoning across turns, and lets you control thinking per request for speed or accuracy. It upgrades agentic workflows with stronger multi‑step tool use, better

Input
Text
Output
Text

GLM-4.7 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for GLM-4.7 across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How GLM-4.7 holds up as conversations get longer.

Quality Tracker

GLM-4.7 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Thu Sep 10 2026
Notice missing or incorrect data?

GLM-4.7 pricing

Providers

GLM-4.7 starts at $0.400 per million input tokens and $1.75 per million output tokens via DeepInfra. Reused prompt prefixes cost $0.0800 per million cached input tokens. See all 3 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.400$0.0800$1.75202.8K/202.8K
/
Fireworks logoFireworks
$0.600$2.20202.8K/131.1K
/
Novita logoNovita
$0.600$2.20204.8K/131.1K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

GLM-4.7 model size

GLM-4.7 has 358 billion parameters. See how it compares to other models in the same parameter range.

Parameters
358BMoE
Frontier (200B+)
358B
1B7B70B405B

GLM-4.7 context window

Input and output token limits for GLM-4.7, plus how it ranks on long-context understanding.

InputOutput
205Ktokens
203Ktokens
308 pages of text
205K
8K128K1M

Try now

huggle
GLM-4.7in Huggle

Make it with
GLM-4.7.

GLM-4.7

GLM-4.7 latency

GLM-4.7 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

GLM-4.7 examples

Recent arena outputs from GLM-4.7, picked from the highest-ranked matchups.

GLM-4.7 license

GLM-4.7 is released under the MIT license, which permits commercial use, has 358.0B parameters.

License
MIT
Commercial use allowed
Parameters
358.0B

MIT License - allows commercial use

GLM-4.7 resources

Official sources for GLM-4.7: provider documentation, official playground, paper or system card, official launch post, model weights.

GLM-4.7 vs other models

The most-compared alternatives to GLM-4.7 are GLM-5.1, GPT-5 High, Seed 2.0 Lite. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like GLM-4.7

Models ranked just above and below GLM-4.7 by LLM Stats score.

 

GLM-5.1

Score pending
 

GPT-5 High

Score pending
 

Seed 2.0 Lite

Score pending
 

K-EXAONE-236B-A23B

Score pending
 

Step-3.5-Flash

Score pending
 

GPT-5 Codex

Score pending

FAQ

Common questions about GLM-4.7.

When was GLM-4.7 released?

GLM-4.7 was released on December 22, 2025 by ZAI. This is the official GLM-4.7 release date tracked on LLM Stats.

How much does GLM-4.7 cost?

GLM-4.7 pricing starts at $0.40 per million input tokens, $0.08 per million cached input tokens, $1.75 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is GLM-4.7?

GLM-4.7 has 358 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created GLM-4.7?

GLM-4.7 was created by ZAI.

What is the license for GLM-4.7?

GLM-4.7 is released under the MIT license. This is an open-source / open-weight license that permits self-hosting.

Is GLM-4.7 multimodal?

Yes, GLM-4.7 is multimodal and can accept both text and images as input.

Where can I use GLM-4.7?

GLM-4.7 is available through 3 providers including DeepInfra, Fireworks, Novita.

Where is the GLM-4.7 paper or technical report?

GLM-4.7 has a paper or technical report available at https://arxiv.org/abs/2508.06471. Use that source for architecture, training, release and evaluation details.

What models should I compare GLM-4.7 against?

Common GLM-4.7 comparisons include GLM-4.7 vs GLM-5.1, GLM-4.7 vs GPT-5 High, GLM-4.7 vs Seed 2.0 Lite. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.