The AI arena is free today

Open Superagent
MistralReleased on Jul 18, 2024

Mistral NeMo Instruct: Benchmarks, Pricing & Context Window

Mistral NeMo Instruct is a language model from Mistral, released in July 2024, with a 131K-token context window, and pricing from $0.019/M input and $0.030/M output.

A state-of-the-art 12B multilingual model with a 128k context window, designed for global applications and strong in multiple languages.

Input
Text
Output
Text

Mistral NeMo Instruct benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Mistral NeMo Instruct across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Mistral NeMo Instruct holds up as conversations get longer.

Quality Tracker

Mistral NeMo Instruct Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Thu Sep 10 2026
Notice missing or incorrect data?

Mistral NeMo Instruct pricing

Providers

Mistral NeMo Instruct starts at $0.0190 per million input tokens and $0.0300 per million output tokens via DeepInfra. See all 3 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.0190$0.0300131.1K/131.1K
/
Google logoGoogle
$0.150$0.150128.0K/128.0K
0.40
/
Mistral AI logoMistral AI
$0.150$0.150128.0K/128.0K
0.50
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Mistral NeMo Instruct model size

Mistral NeMo Instruct has 12 billion parameters. See how it compares to other models in the same parameter range.

Parameters
12B
Medium (10–30B)
12B
1B7B70B405B

Mistral NeMo Instruct context window

Input and output token limits for Mistral NeMo Instruct, plus how it ranks on long-context understanding.

InputOutput
131Ktokens
131Ktokens
197 pages of text
131K
8K128K1M

Try now

huggle
Mistral NeMo Instructin Huggle

Make it with
Mistral NeMo Instruct.

Mistral NeMo Instruct

Mistral NeMo Instruct latency

Mistral NeMo Instruct time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Mistral NeMo Instruct examples

Recent arena outputs from Mistral NeMo Instruct, picked from the highest-ranked matchups.

Mistral NeMo Instruct license

Mistral NeMo Instruct is released under the Apache 2.0 license, which permits commercial use, has 12.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
12.0B

Apache License 2.0 - allows commercial use

Mistral NeMo Instruct resources

Official sources for Mistral NeMo Instruct: provider documentation, official launch post, source repository.

Mistral NeMo Instruct vs other models

The most-compared alternatives to Mistral NeMo Instruct are Qwen2.5 32B Instruct, Phi 4 Mini, Granite 3.3 8B Base. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Mistral NeMo Instruct

Models ranked just above and below Mistral NeMo Instruct by LLM Stats score.

 

Qwen2.5 32B Instruct

Score pending
 

Phi 4 Mini

Score pending
 

Granite 3.3 8B Base

Score pending
 

Phi-3.5-MoE-instruct

Score pending
 

Gemma 2 9B

Score pending
 

Qwen2.5-Coder 32B Instruct

Score pending

FAQ

Common questions about Mistral NeMo Instruct.

When was Mistral NeMo Instruct released?

Mistral NeMo Instruct was released on July 18, 2024 by Mistral. This is the official Mistral NeMo Instruct release date tracked on LLM Stats.

How much does Mistral NeMo Instruct cost?

Mistral NeMo Instruct pricing starts at $0.02 per million input tokens and $0.03 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is Mistral NeMo Instruct?

Mistral NeMo Instruct has 12 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Mistral NeMo Instruct?

Mistral NeMo Instruct was created by Mistral.

What is the license for Mistral NeMo Instruct?

Mistral NeMo Instruct is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

What is Mistral NeMo Instruct latency?

Mistral NeMo Instruct p95 time to first token is 0.40 seconds via Google over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use Mistral NeMo Instruct?

Mistral NeMo Instruct is available through 3 providers including DeepInfra, Google, Mistral AI.

Where is the Mistral NeMo Instruct paper or technical report?

Mistral NeMo Instruct has a paper or technical report available at https://mistral.ai/news/mistral-nemo/. Use that source for architecture, training, release and evaluation details.

What models should I compare Mistral NeMo Instruct against?

Common Mistral NeMo Instruct comparisons include Mistral NeMo Instruct vs Qwen2.5 32B Instruct, Mistral NeMo Instruct vs Phi 4 Mini, Mistral NeMo Instruct vs Granite 3.3 8B Base. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.