LLMcompare

Search

Search for a command to run...

A · open · GLM

GLM-5.1

Zhipu AI

B · open · GLM

GLM-5.3

Zhipu AI

0:11

overall capability wins (2 ties across 13 scored). GLM-5.3 leads overall.

  • Reasoning0:5
  • Coding0:3
  • Arena0:1
  • Tool use0:1
  • Composite indices0:2
  • Specs0:1

Verdict

GLM-5.3 leads coding (widest gap: +106.1 on WebDev Elo); GLM-5.3 ranks higher on LMArena (+18 Elo); GLM-5.3 offers ~4.9x larger context (1M vs 203K).

  • GLM-5.3 leads coding (widest gap: +106.1 on WebDev Elo)
  • GLM-5.3 ranks higher on LMArena (+18 Elo)
  • GLM-5.3 offers ~4.9x larger context (1M vs 203K)

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

GLM-5.1
$2.50
GLM-5.3
-

Price per 1M tokens

Full comparison

Organization

Zhipu AI

Zhipu AI

Family

GLM

GLM

License

MIT

GLM-5.3 License

Open weights

Yes

Yes

Release

Apr 7, 2026

Aug 14, 2026

Knowledge cutoff

-

-

API / provider

Zhipu / Z.ai

No primary API price listed

Modalities

text → text

text → text

Specs

Context window

203K

1M

B +797K

Max output

128K

128K

tie

Parameters

744B (40B act.)

744B (40B act.)

tie

Pricing

Input $/1M

$1.40

—

—

Output $/1M

$4.40

—

—

Blended $/1M (3∶1)

$2.15

-

—

Speed

tok/s

-

-

—

TTFT (s)

-

-

—

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

86.8

91.7

B +4.9 pts

Humanity's Last Exam

30.1

42.3

B +12.2 pts

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

23.7

33.9

B +10.2 pts

AA-LCR v1.1

73.7

79.7

B +6.0 pts

CritPt

4.6

19.1

B +14.5 pts

MMMU-Pro

—

—

—

IFBench

76.3

—

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

61.8

83.9

B +22.1 pts

Aider Polyglot

—

—

—

Terminal-Bench 3

—

28.3

—

BigCodeBench

—

—

—

SciCode

44.8

59.0

B +14.2 pts

CursorBench

—

—

—

SWE-Rebench

—

—

—

NL2Repo-Bench

—

58.0

—

DeepSWE

—

66.9

—

WebDev Arena

1,508

1,614

B +106 Elo

Terminal-Bench 4.0

—

41.8

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

1,466

1,483

B +18 Elo

GLM-5.1

Context
203K / 128K out
Parameters
744B (40B act.)
Price
$1.40 / $4.40
Speed
—
Modalities
text → text
License
MIT

Zhipu's April 2026 refinement of the GLM-5 744B MoE (40B active) flagship for agentic engineering — MIT open weights, ~200K context, 128K max output. Benchmark figures here are Artificial Analysis and LMArena leaderboard snapshots, not Zhipu's own launch numbers.

GLM-5.3

Context
1M / 128K out
Parameters
744B (40B act.)
Price
—
Speed
—
Modalities
text → text
License
GLM-5.3 License

Zhipu's Aug 14, 2026 flagship — GLM-5.2 base re-post-trained for complex software engineering, terminal, and real-world agent tasks, with SOTA open-weights coding and emergent cybersecurity skills (CyberGym 84.5). Open weights (744B MoE / 40B active) are on Hugging Face under the GLM-5.3 license; thinking effort is low/high/max (default max).

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

GPQA: B +4.9HLE: B +12.2CritPt: B +14.5AA-LCR: B +6.0Omniscience Acc.: B +10.2

Coding

TermBench: B +22.1SciCode: B +14.2WebDev Elo: B +106

Tool use & function calling

τ³-Banking: B +36.7

Arena

Arena Elo: B +18

Composite indices

AA Index: B +22.2AA Index v4.3: B +18.5

Pick a different pair · Back to catalog