LLMcompare

Search

Search for a command to run...

A · open · GLM

GLM-5.1

Zhipu AI

B · open · GLM

GLM-5.2

Zhipu AI

3:10

overall capability wins (0 ties across 13 scored). GLM-5.2 leads overall.

  • Reasoning1:5
  • Coding0:3
  • Arena0:1
  • Tool use0:1
  • Composite indices0:2
  • Specs2:0

Verdict

GLM-5.2 leads coding (widest gap: +83.5 on WebDev Elo); GLM-5.2 ranks higher on LMArena (+7 Elo); GLM-5.1 offers larger context (203K vs 200K).

  • GLM-5.2 leads coding (widest gap: +83.5 on WebDev Elo)
  • GLM-5.2 ranks higher on LMArena (+7 Elo)
  • GLM-5.1 offers larger context (203K vs 200K)
  • GLM-5.1 allows ~3.9x more max output (128K vs 33K)

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

GLM-5.1
$2.50
GLM-5.2
$2.50

Price per 1M tokens

Full comparison

Organization

Zhipu AI

Zhipu AI

Family

GLM

GLM

License

MIT

MIT

Open weights

Yes

Yes

Release

Apr 7, 2026

Jun 1, 2026

Knowledge cutoff

-

-

API / provider

Zhipu / Z.ai

Zhipu / Z.ai

Modalities

text → text

text, image → text

Specs

Context window

203K

200K

A +3K

Max output

128K

33K

A +95K

Parameters

744B (40B act.)

—

—

Pricing

Input $/1M

$1.40

$1.40

tie

Output $/1M

$4.40

$4.40

tie

Blended $/1M (3∶1)

$2.15

$2.15

tie

Speed

tok/s

-

168

—

TTFT (s)

-

0.35

—

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

86.8

89.5

B +2.7 pts

Humanity's Last Exam

30.1

41.1

B +11.0 pts

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

23.7

24.3

B +0.6 pts

AA-LCR v1.1

73.7

78.3

B +4.6 pts

CritPt

4.6

20.9

B +16.3 pts

MMMU-Pro

—

—

—

IFBench

76.3

73.3

A +3.0 pts

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

61.8

77.9

B +16.1 pts

Aider Polyglot

—

—

—

Terminal-Bench 3

—

4.6

—

BigCodeBench

—

—

—

SciCode

44.8

51.2

B +6.4 pts

CursorBench

—

55.0

—

SWE-Rebench

—

57.0

—

NL2Repo-Bench

—

—

—

DeepSWE

—

44.0

—

WebDev Arena

1,508

1,592

B +84 Elo

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

1,466

1,472

B +7 Elo

GLM-5.1

Context
203K / 128K out
Parameters
744B (40B act.)
Price
$1.40 / $4.40
Speed
—
Modalities
text → text
License
MIT

Zhipu's April 2026 refinement of the GLM-5 744B MoE (40B active) flagship for agentic engineering — MIT open weights, ~200K context, 128K max output. Benchmark figures here are Artificial Analysis and LMArena leaderboard snapshots, not Zhipu's own launch numbers.

GLM-5.2

Context
200K / 33K out
Parameters
—
Price
$1.40 / $4.40
Speed
168 tok/s · 0.35s TTFT
Modalities
text, image → text
License
MIT

Open MIT GLM-5.2 - competitive open Arena Elo and strong Chinese/English bilingual performance.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

GPQA: B +2.7HLE: B +11.0CritPt: B +16.3AA-LCR: B +4.6Omniscience Acc.: B +0.6IFBench: A +3.0

Coding

TermBench: B +16.1SciCode: B +6.4WebDev Elo: B +84

Tool use & function calling

τ³-Banking: B +21.0

Arena

Arena Elo: B +7

Composite indices

AA Index: B +7.6AA Index v4.3: B +7.6

Pick a different pair · Back to catalog