vs
2:11
overall capability wins (1 ties across 14 scored). Mistral Medium 3.1 leads overall.
- Reasoning1:7
- Coding0:3
- Tool use0:1
- Composite indices0:1
- Pricing3:0
- Speed2:0
- Specs1:0
Verdict
Mistral Medium 3.1 leads coding (widest gap: +17.1 on SciCode); Ministral 3 3B offers ~2.0x larger context (256K vs 128K); Ministral 3 3B ships open weights (Apache 2.0).
- Mistral Medium 3.1 leads coding (widest gap: +17.1 on SciCode)
- Ministral 3 3B offers ~2.0x larger context (256K vs 128K)
- Ministral 3 3B ships open weights (Apache 2.0)
- Ministral 3 3B is ~8.0x cheaper on a blended token basis than Mistral Medium 3.1
Cost for 1M in + 250K out
Illustrative chat workload at primary-provider list prices.
- Ministral 3 3B
- $0.13
- Mistral Medium 3.1
- $0.90
Ministral 3 3B is about 7.2x cheaper on this workload.
Price per 1M tokens
Full comparison
Organization
Mistral
Mistral
Family
Mistral
Mistral
License
Apache 2.0
Proprietary
Open weights
Yes
No
Release
Dec 1, 2025
Aug 1, 2025
Knowledge cutoff
-
-
API / provider
Mistral
Mistral
Modalities
text, image → text
text, image → text
Specs
Context window
256K
128K
Max output
33K
33K
Parameters
3B
—
Pricing
Input $/1M
$0.10
$0.40
Output $/1M
$0.10
$2
Blended $/1M (3∶1)
$0.10
$0.80
Speed
tok/s
180
100
TTFT (s)
0.15
0.3
Reasoning
MMLU-Pro
52.4
68.3
GPQA Diamond
35.8
58.8
Humanity's Last Exam
5.4
4.7
AIME 2025
22.0
38.3
MATH-500
—
—
Humanity's Last Exam (with tools)
—
—
AA-Omniscience Accuracy
9.0
20.8
AA-LCR v1.1
17.0
42.7
CritPt
—
—
MMMU-Pro
38.1
54.2
IFBench
26.8
39.8
Chartography
—
—
Chartography (With Tools)
—
—
Coding
SWE-bench Verified
—
—
SWE-bench Pro
—
—
SWE-bench Multilingual
—
—
LiveCodeBench
24.7
40.6
Terminal-Bench 2.1
0.0
13.9
Aider Polyglot
—
—
Terminal-Bench 3
—
—
BigCodeBench
—
—
SciCode
15.3
32.4
CursorBench
—
—
SWE-Rebench
—
—
NL2Repo-Bench
—
—
DeepSWE
—
—
WebDev Arena
—
—
Terminal-Bench 4.0
—
—
CursorBench 4.0
—
—
FrontierCode v1.1 (Main)
—
—
Arena
LMArena Elo
—
—
| Metric | Ministral 3 3B | Mistral Medium 3.1 | Delta |
|---|---|---|---|
| Identity | |||
| Organization | Mistral | Mistral | — |
| Family | Mistral | Mistral | — |
| License | Apache 2.0 | Proprietary | — |
| Open weights | Yes | No | — |
| Release | Dec 1, 2025 | Aug 1, 2025 | — |
| Knowledge cutoff | - | - | — |
| API / provider | Mistral | Mistral | — |
| Modalities | text, image → text | text, image → text | — |
| Specs | |||
| Context window | 256K | 128K | A +128K |
| Max output | 33K | 33K | tie |
| Parameters | 3B | — | — |
| Pricing | |||
| Input $/1M | $0.10 | $0.40 | A +$0.30 |
| Output $/1M | $0.10 | $2 | A +$1.90 |
| Blended $/1M (3∶1) | $0.10 | $0.80 | A +$0.70 |
| Speed | |||
| tok/s | 180 | 100 | A +80 |
| TTFT (s) | 0.15 | 0.3 | A +0.15 |
| Reasoning | |||
| MMLU-Pro | 52.4 | 68.3 | B +15.9 pts |
| GPQA Diamond | 35.8 | 58.8 | B +23.0 pts |
| Humanity's Last Exam | 5.4 | 4.7 | A +0.7 pts |
| AIME 2025 | 22.0 | 38.3 | B +16.3 pts |
| MATH-500 | — | — | — |
| Humanity's Last Exam (with tools) | — | — | — |
| AA-Omniscience Accuracy | 9.0 | 20.8 | B +11.8 pts |
| AA-LCR v1.1 | 17.0 | 42.7 | B +25.7 pts |
| CritPt | — | — | — |
| MMMU-Pro | 38.1 | 54.2 | B +16.1 pts |
| IFBench | 26.8 | 39.8 | B +13.0 pts |
| Chartography | — | — | — |
| Chartography (With Tools) | — | — | — |
| Coding | |||
| SWE-bench Verified | — | — | — |
| SWE-bench Pro | — | — | — |
| SWE-bench Multilingual | — | — | — |
| LiveCodeBench | 24.7 | 40.6 | B +15.9 pts |
| Terminal-Bench 2.1 | 0.0 | 13.9 | B +13.9 pts |
| Aider Polyglot | — | — | — |
| Terminal-Bench 3 | — | — | — |
| BigCodeBench | — | — | — |
| SciCode | 15.3 | 32.4 | B +17.1 pts |
| CursorBench | — | — | — |
| SWE-Rebench | — | — | — |
| NL2Repo-Bench | — | — | — |
| DeepSWE | — | — | — |
| WebDev Arena | — | — | — |
| Terminal-Bench 4.0 | — | — | — |
| CursorBench 4.0 | — | — | — |
| FrontierCode v1.1 (Main) | — | — | — |
| Arena | |||
| LMArena Elo | — | — | — |
Ministral 3 3B
- Context
- 256K / 33K out
- Parameters
- 3B
- Price
- $0.10 / $0.10
- Speed
- 180 tok/s · 0.15s TTFT
- Modalities
- text, image → text
- License
- Apache 2.0
Tiny Ministral 3 multimodal model - cheapest Mistral hosted tier for edge and classification.
Mistral Medium 3.1
- Context
- 128K / 33K out
- Parameters
- —
- Price
- $0.40 / $2
- Speed
- 100 tok/s · 0.3s TTFT
- Modalities
- text, image → text
- License
- Proprietary
Cost-efficient Mistral mid-tier for enterprise chat and document tasks.
Benchmark charts
Winner bars are emphasized. Per-benchmark deltas sit above each chart.