The AI arena is free today

Open Superagent
MiniMaxReleased on Oct 27, 2025

MiniMax M2: Benchmarks, Pricing & Context Window

MiniMax M2 is a language model from MiniMax, released in October 2025, with a 1M-token context window, and pricing from $0.300/M input and $1.20/M output.

MiniMax M2 is an open-source large language model by MiniMax, built for agents and coding tasks. It delivers state-of-the-art tool use, reasoning, and search performance while maintaining exceptional cost-efficiency and speed, priced at

Input
Text
Output
Text

MiniMax M2 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for MiniMax M2 across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How MiniMax M2 holds up as conversations get longer.

Quality Tracker

MiniMax M2 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Sep 14 2026
Notice missing or incorrect data?

MiniMax M2 pricing

Providers

MiniMax M2 starts at $0.300 per million input tokens and $1.20 per million output tokens via MiniMax. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
MiniMax logoMiniMax
$0.300$1.201.0M/1.0M
4.00
/
Novita logoNovita
$0.300$0.0300$1.20204.8K/131.1K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

MiniMax M2 model size

MiniMax M2 has 230 billion parameters. See how it compares to other models in the same parameter range.

Parameters
230BMoE
Frontier (200B+)
230B
1B7B70B405B

MiniMax M2 context window

Input and output token limits for MiniMax M2, plus how it ranks on long-context understanding.

InputOutput
1Mtokens
1Mtokens
1.5k pages of text
1M
8K128K1M

Try now

huggle
MiniMax M2in Huggle

Make it with
MiniMax M2.

MiniMax M2

MiniMax M2 latency

MiniMax M2 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

MiniMax M2 examples

Recent arena outputs from MiniMax M2, picked from the highest-ranked matchups.

MiniMax M2 license

MiniMax M2 is released under the MIT license, which permits commercial use, has 230.0B parameters.

License
MIT
Commercial use allowed
Parameters
230.0B

MIT License - allows commercial use

MiniMax M2 resources

Official sources for MiniMax M2: provider documentation, official launch post, source repository, model weights.

MiniMax M2 vs other models

The most-compared alternatives to MiniMax M2 are Phi 4 Reasoning Plus, Qwen3 VL 235B A22B Instruct, Qwen3 VL 32B Thinking. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like MiniMax M2

Models ranked just above and below MiniMax M2 by LLM Stats score.

 

Phi 4 Reasoning Plus

Score pending
 

Qwen3 VL 235B A22B Instruct

Score pending
 

Qwen3 VL 32B Thinking

Score pending
 

Sarvam-105B

Score pending
 

Claude Opus 4.1

Score pending
 

Qwen3-235B-A22B-Instruct-2507

Score pending

FAQ

Common questions about MiniMax M2.

When was MiniMax M2 released?

MiniMax M2 was released on October 27, 2025 by MiniMax. This is the official MiniMax M2 release date tracked on LLM Stats.

How much does MiniMax M2 cost?

MiniMax M2 pricing starts at $0.30 per million input tokens and $1.20 per million output tokens via MiniMax, the lowest price among tracked providers.

How big is MiniMax M2?

MiniMax M2 has 230 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created MiniMax M2?

MiniMax M2 was created by MiniMax.

What is the license for MiniMax M2?

MiniMax M2 is released under the MIT license. This is an open-source / open-weight license that permits self-hosting.

What is MiniMax M2 latency?

MiniMax M2 p95 time to first token is 4.00 seconds via MiniMax over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use MiniMax M2?

MiniMax M2 is available through 2 providers including MiniMax, Novita.

Where is the MiniMax M2 paper or technical report?

MiniMax M2 has a paper or technical report available at https://www.minimax.io/news/minimax-m2. Use that source for architecture, training, release and evaluation details.

What models should I compare MiniMax M2 against?

Common MiniMax M2 comparisons include MiniMax M2 vs Phi 4 Reasoning Plus, MiniMax M2 vs Qwen3 VL 235B A22B Instruct, MiniMax M2 vs Qwen3 VL 32B Thinking. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.