LLMcompare

Search

Search for a command to run...

A · open · Nemotron

Nemotron 3 Ultra

NVIDIA

B · open · Nemotron

Nemotron Nano 3 30B

NVIDIA

4:1

overall capability wins (1 ties across 6 scored). Nemotron 3 Ultra leads overall.

  • Reasoning2:0
  • Coding1:0
  • Pricing0:3
  • Speed2:0
  • Specs1:1

Verdict

Nemotron 3 Ultra leads coding (widest gap: +20.7 on LiveCode); Nemotron Nano 3 30B allows ~2.0x more max output (66K vs 33K); Nemotron Nano 3 30B is ~10.3x cheaper on a blended token basis than Nemotron 3 Ultra.

  • Nemotron 3 Ultra leads coding (widest gap: +20.7 on LiveCode)
  • Nemotron Nano 3 30B allows ~2.0x more max output (66K vs 33K)
  • Nemotron Nano 3 30B is ~10.3x cheaper on a blended token basis than Nemotron 3 Ultra

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Nemotron 3 Ultra
$1.05
Nemotron Nano 3 30B
$0.10

Nemotron Nano 3 30B is about 10.5x cheaper on this workload.

Price per 1M tokens

Full comparison

Organization

NVIDIA

NVIDIA

Family

Nemotron

Nemotron

License

NVIDIA Open Model

NVIDIA Open Model

Open weights

Yes

Yes

Release

Jan 15, 2026

Dec 15, 2025

Knowledge cutoff

-

-

API / provider

NVIDIA NIM / self-host

NVIDIA NIM

Modalities

text → text

text → text

Specs

Context window

262K

262K

tie

Max output

33K

66K

B +33K

Parameters

550B (55B act.)

30B (3.5B act.)

A +520

Pricing

Input $/1M

$0.60

$0.05

B +$0.55

Output $/1M

$1.80

$0.20

B +$1.60

Blended $/1M (3∶1)

$0.90

$0.09

B +$0.81

Speed

tok/s

100

68

A +32

TTFT (s)

0.3

0.4

A +0.10

Reasoning

MMLU-Pro

86.8

78.3

A +8.5 pts

GPQA Diamond

86.7

—

—

Humanity's Last Exam

28.4

10.6

A +17.8 pts

AIME 2025

—

89.1

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

22.6

—

—

AA-LCR v1.1

79.3

—

—

CritPt

3.1

—

—

MMMU-Pro

—

—

—

IFBench

81.4

—

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

70.7

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

67.7

—

—

LiveCodeBench

89.0

68.3

A +20.7 pts

Terminal-Bench 2.1

56.4

—

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

—

—

BigCodeBench

—

—

—

SciCode

40.3

—

—

CursorBench

—

—

—

SWE-Rebench

—

—

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

—

—

—

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

—

—

—

Nemotron 3 Ultra

Context
262K / 33K out
Parameters
550B (55B act.)
Price
$0.60 / $1.80
Speed
100 tok/s · 0.3s TTFT
Modalities
text → text
License
NVIDIA Open Model

NVIDIA open MoE optimized for NIM deployment and enterprise self-hosting.

Nemotron Nano 3 30B

Context
262K / 66K out
Parameters
30B (3.5B act.)
Price
$0.05 / $0.20
Speed
68 tok/s · 0.4s TTFT
Modalities
text → text
License
NVIDIA Open Model

Efficient Nemotron Nano 3 MoE — Mamba-2 + attention hybrid with 3.5B active params and 262K context on NIM.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

MMLU-Pro: A +8.5HLE: A +17.8

Coding

LiveCode: A +20.7

Pick a different pair · Back to catalog