The AI arena is free today

Open Superagent

Devstral Small 1.1 vs Ministral 3 (14B Base 2512)

Devstral Small 1.1 and Ministral 3 (14B Base 2512) are closely matched at 9.3 and 3.6 on the LLM Stats Score.

Mistral AI · Mistral AI · Updated for 2026

Which is better?

Devstral Small 1.1 and Ministral 3 (14B Base 2512) are closely matched on the overall LLM Stats Score at 9.3 and 3.6.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Devstral Small 1.1

  • you want predictable pricing at $0.10/M input and $0.30/M output

Choose Ministral 3 (14B Base 2512)

  • you want the most recent training data — it shipped Dec 2025

At a glance

The differences that matter most.

Core performance indexes
9.3
#275
3.6
#314
9.2
#270
3.5
#306
Cost, coverage & limits
Benchmark wins
—
—
Input price
$0.10 / M
— / M
Output price
$0.30 / M
— / M
Context window
128,000
—

Individual benchmarks

1 reported for Devstral Small 1.1 · 6 for Ministral 3 (14B Base 2512)

No common benchmarks found

Devstral Small 1.1 and Ministral 3 (14B Base 2512)don't have any common benchmark datasets to compare. They may have been evaluated on different testing suites.

Human preference

Blind head-to-head votes and playground preference scores

Model Size

Parameter count comparison

10.0B diff

Devstral Small 1.1 has 10.0B more parameters than Ministral 3 (14B Base 2512), making it 71.4% larger.

Mistral AI
Devstral Small 1.1
24.0Bparameters
Mistral AI
Ministral 3 (14B Base 2512)
14.0Bparameters
24.0B
Devstral Small 1.1
14.0B
Ministral 3 (14B Base 2512)

Context Window

Maximum input and output token capacity

Only Devstral Small 1.1 specifies input context (128,000 tokens). Only Devstral Small 1.1 specifies output context (128,000 tokens).

Mistral AI
Devstral Small 1.1
Input128,000 tokens
Output128,000 tokens
Mistral AI
Ministral 3 (14B Base 2512)
Input- tokens
Output- tokens
Mon Sep 28 2026 • llm-stats.com

Input capabilities

Documented input modalities across available providers

Ministral 3 (14B Base 2512) supports multimodal inputs, whereas Devstral Small 1.1 does not.

Ministral 3 (14B Base 2512) can handle both text and other forms of data like images, making it suitable for multimodal applications.

Devstral Small 1.1

Text
Images
Audio
Video

Ministral 3 (14B Base 2512)

Text
Images
Audio
Video

License

Usage and distribution terms

Both models are licensed under Apache 2.0.

Both models share the same licensing terms, providing consistent usage rights.

Devstral Small 1.1

Apache 2.0

Open weights

Ministral 3 (14B Base 2512)

Apache 2.0

Open weights

Release Timeline

When each model was launched

Devstral Small 1.1 was released on 2025-07-11, while Ministral 3 (14B Base 2512) was released on 2025-12-04.

Ministral 3 (14B Base 2512) is 5 months newer than Devstral Small 1.1.

Devstral Small 1.1

Jul 11, 2025

1.2 years ago

Ministral 3 (14B Base 2512)

Dec 4, 2025

9 months ago

4mo newer

Knowledge Cutoff

When training data ends

Neither model specifies a knowledge cutoff date.

Unable to compare the recency of their training data.

No cutoff dates available

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion→

Judge for yourself.

Run your own prompts against Devstral Small 1.1 and Ministral 3 (14B Base 2512) side-by-side, then vote on the output you prefer.

Devstral Small 1.1
✓ Preferred
Ministral 3 (14B Base 2512)
Open in Playground

FAQ

Common questions about Devstral Small 1.1 vs Ministral 3 (14B Base 2512).

Which is better, Devstral Small 1.1 or Ministral 3 (14B Base 2512)?

Devstral Small 1.1 and Ministral 3 (14B Base 2512) are closely matched on the LLM Stats Score at 9.3 and 3.6. Devstral Small 1.1 is made by Mistral AI and Ministral 3 (14B Base 2512) is made by Mistral AI. The best choice depends on your use case — compare their capability indexes, individual benchmarks, pricing, and limits above.

How does Devstral Small 1.1 compare to Ministral 3 (14B Base 2512) in benchmarks?

Devstral Small 1.1 scores SWE-Bench Verified: 53.6%. Ministral 3 (14B Base 2512) scores MMLU-Redux: 82.0%, MMLU: 79.4%, TriviaQA: 74.9%, Multilingual MMLU: 74.2%, MATH (CoT): 67.6%.

What are the context window sizes for Devstral Small 1.1 and Ministral 3 (14B Base 2512)?

Devstral Small 1.1 supports 128K tokens and Ministral 3 (14B Base 2512) supports an unknown number of tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Devstral Small 1.1 and Ministral 3 (14B Base 2512)?

Key differences include LLM Stats Score (9.3 vs 3.6), multimodal support (no vs yes). See the full comparison above for benchmark-by-benchmark results.