LLMcompare

Search

Search for a command to run...

A · open · Mistral

Ministral 3 14B

Mistral

B · closed · Mistral

Mistral Large 2

Mistral

5:3

overall capability wins (0 ties across 8 scored). Ministral 3 14B leads overall.

  • Reasoning3:1
  • Coding0:1
  • Pricing3:0
  • Speed2:0
  • Specs2:1

Verdict

Mistral Large 2 leads coding (widest gap: +5.4 on SciCode); Ministral 3 14B edges reasoning & knowledge; Ministral 3 14B offers ~2.0x larger context (256K vs 128K).

  • Mistral Large 2 leads coding (widest gap: +5.4 on SciCode)
  • Ministral 3 14B edges reasoning & knowledge
  • Ministral 3 14B offers ~2.0x larger context (256K vs 128K)
  • Ministral 3 14B allows ~4.0x more max output (33K vs 8K)
  • Ministral 3 14B ships open weights (Apache 2.0)
  • Ministral 3 14B is ~15.0x cheaper on a blended token basis than Mistral Large 2

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Ministral 3 14B
$0.25
Mistral Large 2
$3.50

Ministral 3 14B is about 14.0x cheaper on this workload.

Price per 1M tokens

Full comparison

Organization

Mistral

Mistral

Family

Mistral

Mistral

License

Apache 2.0

Mistral Research / Proprietary API

Open weights

Yes

No

Release

Dec 1, 2025

Jul 24, 2024

Knowledge cutoff

-

-

API / provider

Mistral

Mistral

Modalities

text, image → text

text → text

Specs

Context window

256K

128K

A +128K

Max output

33K

8K

A +25K

Parameters

14B

123B

B +109

Pricing

Input $/1M

$0.20

$2

A +$1.80

Output $/1M

$0.20

$6

A +$5.80

Blended $/1M (3∶1)

$0.20

$3

A +$2.80

Speed

tok/s

130

65

A +65

TTFT (s)

0.2

0.45

A +0.25

Reasoning

MMLU-Pro

69.3

—

—

GPQA Diamond

57.2

48.6

A +8.6 pts

Humanity's Last Exam

4.6

3.3

A +1.3 pts

AIME 2025

30.0

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

13.6

19.9

B +6.3 pts

AA-LCR v1.1

26.3

—

—

CritPt

—

—

—

MMMU-Pro

49.8

—

—

IFBench

32.0

31.2

A +0.8 pts

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

35.1

—

—

Terminal-Bench 2.1

9.7

—

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

—

—

BigCodeBench

—

—

—

SciCode

23.8

29.2

B +5.4 pts

CursorBench

—

—

—

SWE-Rebench

—

—

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

—

—

—

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

—

—

—

Ministral 3 14B

Context
256K / 33K out
Parameters
14B
Price
$0.20 / $0.20
Speed
130 tok/s · 0.2s TTFT
Modalities
text, image → text
License
Apache 2.0

Largest Ministral 3 edge multimodal model with symmetric $0.20/$0.20 pricing.

Mistral Large 2

Context
128K / 8K out
Parameters
123B
Price
$2 / $6
Speed
65 tok/s · 0.45s TTFT
Modalities
text → text
License
Mistral Research / Proprietary API

123B flagship before Large 3 — EU-hosted frontier option with strong multilingual coding.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

GPQA: A +8.6HLE: A +1.3Omniscience Acc.: B +6.3IFBench: A +0.8

Coding

SciCode: B +5.4

Pick a different pair · Back to catalog