LLMcompare

Search

Search for a command to run...

A · open · Granite

Granite 4.1 8B

IBM

B · open · Granite

Granite 4.0 Micro

IBM

6:2

overall capability wins (0 ties across 8 scored). Granite 4.1 8B leads overall.

  • Reasoning4:1
  • Coding1:0
  • Pricing3:0
  • Speed1:1
  • Specs1:1

Verdict

Granite 4.1 8B leads coding (widest gap: +9.9 on SciCode); Granite 4.1 8B offers larger context (131K vs 131K); Granite 4.0 Micro allows ~16.0x more max output (131K vs 8K).

  • Granite 4.1 8B leads coding (widest gap: +9.9 on SciCode)
  • Granite 4.1 8B offers larger context (131K vs 131K)
  • Granite 4.0 Micro allows ~16.0x more max output (131K vs 8K)
  • Granite 4.1 8B is ~1.8x cheaper on a blended token basis than Granite 4.0 Micro

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Granite 4.1 8B
$0.08
Granite 4.0 Micro
$0.13

Granite 4.1 8B is about 1.7x cheaper on this workload.

Price per 1M tokens

Full comparison

Organization

IBM

IBM

Family

Granite

Granite

License

Apache 2.0

Apache 2.0

Open weights

Yes

Yes

Release

Apr 30, 2026

Oct 20, 2025

Knowledge cutoff

-

-

API / provider

IBM watsonx

IBM watsonx

Modalities

text → text

text → text

Specs

Context window

131K

131K

A +72

Max output

8K

131K

B +123K

Parameters

8B

—

—

Pricing

Input $/1M

$0.05

$0.06

A +$0.01

Output $/1M

$0.10

$0.27

A +$0.17

Blended $/1M (3∶1)

$0.06

$0.11

A +$0.05

Speed

tok/s

117

130

B +13

TTFT (s)

0.13

0.15

A +0.02

Reasoning

MMLU-Pro

—

44.7

—

GPQA Diamond

43.3

33.6

A +9.7 pts

Humanity's Last Exam

3.8

5.0

B +1.2 pts

AIME 2025

—

6.0

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

12.3

9.3

A +3.0 pts

AA-LCR v1.1

13.3

7.0

A +6.3 pts

CritPt

—

—

—

MMMU-Pro

—

—

—

IFBench

38.6

24.8

A +13.8 pts

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

18.0

—

Terminal-Bench 2.1

3.4

—

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

—

—

BigCodeBench

—

—

—

SciCode

21.8

11.9

A +9.9 pts

CursorBench

—

—

—

SWE-Rebench

—

—

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

1,192

—

—

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

1,306

—

—

Granite 4.1 8B

Context
131K / 8K out
Parameters
8B
Price
$0.05 / $0.10
Speed
117 tok/s · 0.13s TTFT
Modalities
text → text
License
Apache 2.0

Token-efficient IBM enterprise 8B with tool calling and fill-in-the-middle code support under Apache 2.0.

Granite 4.0 Micro

Context
131K / 131K out
Parameters
—
Price
$0.06 / $0.27
Speed
130 tok/s · 0.15s TTFT
Modalities
text → text
License
Apache 2.0

Ultra-compact Granite 4 hybrid MoE for edge and high-volume watsonx pipelines — among the cheapest hosted enterprise LLMs.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

GPQA: A +9.7HLE: B +1.2AA-LCR: A +6.3Omniscience Acc.: A +3.0IFBench: A +13.8

Coding

SciCode: A +9.9

Pick a different pair · Back to catalog