The AI arena is free today

Open Superagent
MistralReleased on Jan 30, 2025

Mistral Small 3 24B Instruct: Benchmarks, Pricing & Context Window

Mistral Small 3 24B Instruct is a language model from Mistral, released in January 2025, with a 33K-token context window, and pricing from $0.050/M input and $0.080/M output.

Mistral Small 3 is a 24B-parameter LLM licensed under Apache-2.0. It focuses on low-latency, high-efficiency instruction following, maintaining performance comparable to larger models. It provides quick, accurate responses for

Input
Text
Output
Text

Mistral Small 3 24B Instruct benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Mistral Small 3 24B Instruct across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Mistral Small 3 24B Instruct holds up as conversations get longer.

Quality Tracker

Mistral Small 3 24B Instruct Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Thu Oct 08 2026
Notice missing or incorrect data?

Mistral Small 3 24B Instruct pricing

Providers

Mistral Small 3 24B Instruct starts at $0.0500 per million input tokens and $0.0800 per million output tokens via DeepInfra. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$0.0500—$0.080032.8K/32.8K
—
—
/
Mistral AI logoMistral AI
$0.100—$0.30032.0K/32.0K
0.20
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Mistral Small 3 24B Instruct model size

Mistral Small 3 24B Instruct has 24 billion parameters. See how it compares to other models in the same parameter range.

Parameters
24B
Medium (10–30B)
24B
1B7B70B405B

Mistral Small 3 24B Instruct context window

Input and output token limits for Mistral Small 3 24B Instruct, plus how it ranks on long-context understanding.

InputOutput
33Ktokens
33Ktokens
≈ 49 pages of text
33K
8K128K1M

Try now

huggle
Mistral Small 3 24B Instructin Huggle

Make it with
Mistral Small 3 24B Instruct.

Mistral Small 3 24B Instruct

Mistral Small 3 24B Instruct latency

Mistral Small 3 24B Instruct time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Mistral Small 3 24B Instruct examples

Recent arena outputs from Mistral Small 3 24B Instruct, picked from the highest-ranked matchups.

Mistral Small 3 24B Instruct license

Mistral Small 3 24B Instruct is released under the Apache 2.0 license, which permits commercial use, has 24.0B parameters, has a knowledge cutoff of October 2023.

License
Apache 2.0
Commercial use allowed
Parameters
24.0B
Knowledge cutoff
October 2023

Apache License 2.0 - allows commercial use

Mistral Small 3 24B Instruct resources

Official sources for Mistral Small 3 24B Instruct: provider documentation, official launch post, model weights.

Mistral Small 3 24B Instruct vs other models

The most-compared alternatives to Mistral Small 3 24B Instruct are Claude 3.5 Sonnet, Phi 4 Reasoning, Llama 3.1 70B Instruct. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Mistral Small 3 24B Instruct

Models ranked just above and below Mistral Small 3 24B Instruct by LLM Stats score.

 

Claude 3.5 Sonnet

Score pending
 

Phi 4 Reasoning

Score pending
 

Llama 3.1 70B Instruct

Score pending
 

Qwen3 VL 8B Thinking

Score pending
 

Qwen3 VL 4B Thinking

Score pending
 

Qwen2.5 14B Instruct

Score pending

FAQ

Common questions about Mistral Small 3 24B Instruct.

When was Mistral Small 3 24B Instruct released?

Mistral Small 3 24B Instruct was released on January 30, 2025 by Mistral. This is the official Mistral Small 3 24B Instruct release date tracked on LLM Stats.

How much does Mistral Small 3 24B Instruct cost?

Mistral Small 3 24B Instruct pricing starts at $0.05 per million input tokens and $0.08 per million output tokens via DeepInfra, the lowest price among tracked providers.

How big is Mistral Small 3 24B Instruct?

Mistral Small 3 24B Instruct has 24 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Mistral Small 3 24B Instruct?

Mistral Small 3 24B Instruct was created by Mistral.

What is the license for Mistral Small 3 24B Instruct?

Mistral Small 3 24B Instruct is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

What is the knowledge cutoff date for Mistral Small 3 24B Instruct?

Mistral Small 3 24B Instruct has a knowledge cutoff of October 2023, meaning it was trained on data up to that point and may not know about events after it.

What is Mistral Small 3 24B Instruct latency?

Mistral Small 3 24B Instruct p95 time to first token is 0.20 seconds via Mistral AI over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use Mistral Small 3 24B Instruct?

Mistral Small 3 24B Instruct is available through 2 providers including DeepInfra, Mistral AI.

Where is the Mistral Small 3 24B Instruct paper or technical report?

Mistral Small 3 24B Instruct has a paper or technical report available at https://mistral.ai/news/mistral-small-3/. Use that source for architecture, training, release and evaluation details.

What models should I compare Mistral Small 3 24B Instruct against?

Common Mistral Small 3 24B Instruct comparisons include Mistral Small 3 24B Instruct vs Claude 3.5 Sonnet, Mistral Small 3 24B Instruct vs Phi 4 Reasoning, Mistral Small 3 24B Instruct vs Llama 3.1 70B Instruct. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.