XiaomiReleased on Dec 16, 2025

MiMo-V2-Flash: API Pricing, Context Window & Benchmarks

MiMo-V2-Flash is a language model from Xiaomi, released in December 2025.

MiMo-V2-Flash is a powerful, efficient, and ultra-fast foundation language model that excels in reasoning, coding, and agentic scenarios. It is a Mixture-of-Experts model with 309B total parameters and 15B active parameters, featuring a

Input
Text
Output
Text

MiMo-V2-Flash benchmarks

Rankings

Quality Tracker

MiMo-V2-Flash Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Aug 04 2026
Notice missing or incorrect data?

MiMo-V2-Flash pricing

Providers

MiMo-V2-Flash starts at $0.100 per million input tokens and $0.300 per million output tokens via Xiaomi.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Xiaomi logoXiaomi
$0.100$0.300256.0K/16.4K
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

MiMo-V2-Flash model size

MiMo-V2-Flash has 309 billion parameters. See how it compares to other models in the same parameter range.

Parameters
309BMoE
Frontier (200B+)
309B
1B7B70B405B

MiMo-V2-Flash context window

Input and output token limits for MiMo-V2-Flash, plus how it ranks on long-context understanding.

InputOutput
256Ktokens
16Ktokens
385 pages of text
256K
8K128K1M

MiMo-V2-Flash API

Available from the model provider

MiMo-V2-Flash has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

MiMo-V2-Flash latency

MiMo-V2-Flash time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

MiMo-V2-Flash examples

Recent arena outputs from MiMo-V2-Flash, picked from the highest-ranked matchups.

MiMo-V2-Flash license

MiMo-V2-Flash is released under the MIT license, which permits commercial use, has 309.0B parameters.

License
MIT
Commercial use allowed
Parameters
309.0B

MIT License - allows commercial use

MiMo-V2-Flash resources

Official sources for MiMo-V2-Flash: api documentation, official playground, official launch post, model weights.

MiMo-V2-Flash vs other models

The most-compared alternatives to MiMo-V2-Flash are GPT-5 High, Grok-3 Mini, Seed 2.0 Lite. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like MiMo-V2-Flash

Models ranked just above and below MiMo-V2-Flash by LLM Stats score.

 

GPT-5 High

Score pending
 

Grok-3 Mini

Score pending
 

Seed 2.0 Lite

Score pending
 

ChatGPT-4o Latest

Score pending
 

GPT-5.1 Instant

Score pending
 

GPT-5.1 Thinking

Score pending

FAQ

Common questions about MiMo-V2-Flash.

When was MiMo-V2-Flash released?

MiMo-V2-Flash was released on December 16, 2025 by Xiaomi. This is the official MiMo-V2-Flash release date tracked on LLM Stats.

How much does MiMo-V2-Flash cost?

MiMo-V2-Flash pricing starts at $0.10 per million input tokens and $0.30 per million output tokens via Xiaomi, the lowest price among tracked providers.

Is MiMo-V2-Flash available via API?

Yes, MiMo-V2-Flash is available via API. See the official documentation for authentication and endpoint details. It is served by 1 provider tracked on LLM Stats.

How big is MiMo-V2-Flash?

MiMo-V2-Flash has 309 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created MiMo-V2-Flash?

MiMo-V2-Flash was created by Xiaomi.

What is the license for MiMo-V2-Flash?

MiMo-V2-Flash is released under the MIT license. This is an open-source / open-weight license that permits self-hosting.

Where can I use MiMo-V2-Flash?

MiMo-V2-Flash is available through 1 provider including Xiaomi.

Where is the MiMo-V2-Flash paper or technical report?

MiMo-V2-Flash has a paper or technical report available at https://mimo.xiaomi.com/blog/mimo-v2-flash. Use that source for architecture, training, release and evaluation details.

What models should I compare MiMo-V2-Flash against?

Common MiMo-V2-Flash comparisons include MiMo-V2-Flash vs GPT-5 High, MiMo-V2-Flash vs Grok-3 Mini, MiMo-V2-Flash vs Seed 2.0 Lite. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.