The AI arena is free today

Open Superagent

GLM-4.6 vs Grok-3

Both models are evenly matched across the benchmarks. GLM-4.6 is 6.6x cheaper per token.

Zhipu AI · xAI · Updated for 2026

Which is better?

GLM-4.6 outperforms in 1 benchmarks (AIME 2025), while Grok-3 is better at 1 benchmark (GPQA). Both models are evenly matched across the benchmarks.

On price, GLM-4.6 is roughly 6.6x cheaper per token on a blended 3:1 input/output basis, which adds up quickly at production volume.

GLM-4.6 also accepts a larger context window (131,072 input tokens), making it the stronger choice for long documents and large codebases.

Based on current benchmark, pricing, and model metadata for 2026.

Choose GLM-4.6

  • cost matters — it's about 6.6x cheaper per token
  • you process long inputs — it offers a 131,072 token context window
  • you want the most recent training data — it shipped Sep 2025
  • you need open weights you can self-host or fine-tune

Choose Grok-3

  • you want predictable pricing at $3.00/M input and $15.00/M output

At a glance

The differences that matter most.

Benchmark wins
1 of 2
1 of 2
Input price
$0.55 / M
$3.00 / M
Output price
$2.00 / M
$15.00 / M
Context window
131,072
128,000
Released
Sep 2025
Feb 2025
License
MIT
Proprietary

Performance Benchmarks

Comparative analysis across standard metrics

2 benchmarks

GLM-4.6 outperforms in 1 benchmarks (AIME 2025), while Grok-3 is better at 1 benchmark (GPQA).

Both models are evenly matched across the benchmarks.

Mon Aug 24 2026 • llm-stats.com

Arena Performance

Playground indexes and blind preference scores

Pricing Analysis

Price comparison per million tokens

GLM-4.6 costs less

For input processing, GLM-4.6 ($0.55/1M tokens) is 5.5x cheaper than Grok-3 ($3.00/1M tokens).

For output processing, GLM-4.6 ($2.00/1M tokens) is 7.5x cheaper than Grok-3 ($15.00/1M tokens).

In conclusion, Grok-3 is more expensive than GLM-4.6.*

* Using a 3:1 ratio of input to output tokens

Lowest available price from all providers
Mon Aug 24 2026 • llm-stats.com
Zhipu AI
GLM-4.6
Input tokens$0.55
Output tokens$2.00
Best providerFireworks
xAI
Grok-3
Input tokens$3.00
Output tokens$15.00
Best providerxAI
Notice missing or incorrect data?Start an Issue

Context Window

Maximum input and output token capacity

GLM-4.6 accepts 131,072 input tokens compared to Grok-3's 128,000 tokens. GLM-4.6 can generate longer responses up to 131,072 tokens, while Grok-3 is limited to 8,000 tokens.

Zhipu AI
GLM-4.6
Input131,072 tokens
Output131,072 tokens
xAI
Grok-3
Input128,000 tokens
Output8,000 tokens
Mon Aug 24 2026 • llm-stats.com

Input Capabilities

Supported data types and modalities

Both GLM-4.6 and Grok-3 support multimodal inputs.

They are both capable of processing various types of data, offering versatility in application.

GLM-4.6

Text
Images
Audio
Video

Grok-3

Text
Images
Audio
Video

License

Usage and distribution terms

GLM-4.6 is licensed under MIT, while Grok-3 uses a proprietary license.

License differences may affect how you can use these models in commercial or open-source projects.

GLM-4.6

MIT

Open weights

Grok-3

Proprietary

Closed source

Release Timeline

When each model was launched

GLM-4.6 was released on 2025-09-30, while Grok-3 was released on 2025-02-17.

GLM-4.6 is 8 months newer than Grok-3.

GLM-4.6

Sep 30, 2025

10 months ago

7mo newer
Grok-3

Feb 17, 2025

1.5 years ago

Knowledge Cutoff

When training data ends

Grok-3 has a documented knowledge cutoff of 2024-11-17, while GLM-4.6's cutoff date is not specified.

We can confirm Grok-3's training data extends to 2024-11-17, but cannot make a direct comparison without GLM-4.6's cutoff date.

GLM-4.6

Grok-3

Nov 2024

Provider Availability

GLM-4.6 is available from Fireworks, DeepInfra. Grok-3 is available from xAI.

GLM-4.6

fireworks logo
Fireworks
Input Price:Input: $0.55/1MOutput Price:Output: $2.19/1M
deepinfra logo
Deepinfra
Input Price:Input: $0.60/1MOutput Price:Output: $2.00/1M

Grok-3

xai logo
xAI
Input Price:Input: $3.00/1MOutput Price:Output: $15.00/1M
* Prices shown are per million tokens

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against GLM-4.6 and Grok-3 side-by-side, then vote on the output you prefer.

GLM-4.6
✓ Preferred
Grok-3
Open in Playground

FAQ

Common questions about GLM-4.6 vs Grok-3.

Which is better, GLM-4.6 or Grok-3?

Both models are evenly matched across the benchmarks. GLM-4.6 is made by Zhipu AI and Grok-3 is made by xAI. The best choice depends on your use case — compare their benchmark scores, pricing, and capabilities above.

How does GLM-4.6 compare to Grok-3 in benchmarks?

GLM-4.6 scores AIME 2025: 93.9%, LiveCodeBench v6: 82.8%, GPQA: 81.0%, SWE-Bench Verified: 68.0%, BrowseComp: 45.1%. Grok-3 scores AIME 2024: 93.3%, AIME 2025: 93.3%, GPQA: 84.6%, LiveCodeBench: 79.4%, MMMU: 78.0%.

Is GLM-4.6 cheaper than Grok-3?

GLM-4.6 is 5.5x cheaper for input tokens. GLM-4.6 costs $0.55/M input and $2.00/M output via fireworks. Grok-3 costs $3.00/M input and $15.00/M output via xai.

What are the context window sizes for GLM-4.6 and Grok-3?

GLM-4.6 supports 131K tokens and Grok-3 supports 128K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between GLM-4.6 and Grok-3?

Key differences include context window (131K vs 128K), input pricing ($0.55 vs $3.00/M), licensing (MIT vs Proprietary). See the full comparison above for benchmark-by-benchmark results.

Who makes GLM-4.6 and Grok-3?

GLM-4.6 is developed by Zhipu AI and Grok-3 is developed by xAI.