LLMcompare

Search

Search for a command to run...

A · open · GLM

GLM-4.7

Zhipu AI

B · open · GLM

GLM-5.3

Zhipu AI

0:12

overall capability wins (0 ties across 12 scored). GLM-5.3 leads overall.

  • Reasoning0:5
  • Coding0:3
  • Arena0:1
  • Tool use0:1
  • Specs0:2

Verdict

GLM-5.3 leads coding (widest gap: +179.7 on WebDev Elo); GLM-5.3 ranks higher on LMArena (+41 Elo); GLM-5.3 offers ~7.8x larger context (1M vs 128K).

  • GLM-5.3 leads coding (widest gap: +179.7 on WebDev Elo)
  • GLM-5.3 ranks higher on LMArena (+41 Elo)
  • GLM-5.3 offers ~7.8x larger context (1M vs 128K)
  • GLM-5.3 allows ~7.8x more max output (128K vs 16K)

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

GLM-4.7
$0.40
GLM-5.3
-

Price per 1M tokens

Full comparison

Organization

Zhipu AI

Zhipu AI

Family

GLM

GLM

License

MIT

GLM-5.3 License

Open weights

Yes

Yes

Release

Dec 1, 2025

Aug 14, 2026

Knowledge cutoff

-

-

API / provider

Zhipu / Z.ai

No primary API price listed

Modalities

text → text

text → text

Specs

Context window

128K

1M

B +872K

Max output

16K

128K

B +112K

Parameters

—

744B (40B act.)

—

Pricing

Input $/1M

$0.20

—

—

Output $/1M

$0.80

—

—

Blended $/1M (3∶1)

$0.35

-

—

Speed

tok/s

90

-

—

TTFT (s)

0.3

-

—

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

85.9

91.7

B +5.8 pts

Humanity's Last Exam

27.4

42.3

B +14.9 pts

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

29.3

33.9

B +4.6 pts

AA-LCR v1.1

71.0

79.7

B +8.7 pts

CritPt

1.7

19.1

B +17.4 pts

MMMU-Pro

—

—

—

IFBench

67.9

—

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

45.3

83.9

B +38.6 pts

Aider Polyglot

—

—

—

Terminal-Bench 3

—

28.3

—

BigCodeBench

—

—

—

SciCode

45.1

59.0

B +13.9 pts

CursorBench

—

—

—

SWE-Rebench

45.5

—

—

NL2Repo-Bench

—

58.0

—

DeepSWE

—

66.9

—

WebDev Arena

1,435

1,614

B +180 Elo

Terminal-Bench 4.0

—

41.8

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

1,442

1,483

B +41 Elo

GLM-4.7

Context
128K / 16K out
Parameters
—
Price
$0.20 / $0.80
Speed
90 tok/s · 0.3s TTFT
Modalities
text → text
License
MIT

Prior GLM-4.x open checkpoint still useful for cost-sensitive bilingual apps.

GLM-5.3

Context
1M / 128K out
Parameters
744B (40B act.)
Price
—
Speed
—
Modalities
text → text
License
GLM-5.3 License

Zhipu's Aug 14, 2026 flagship — GLM-5.2 base re-post-trained for complex software engineering, terminal, and real-world agent tasks, with SOTA open-weights coding and emergent cybersecurity skills (CyberGym 84.5). Open weights (744B MoE / 40B active) are on Hugging Face under the GLM-5.3 license; thinking effort is low/high/max (default max).

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Reasoning & knowledge

GPQA: B +5.8HLE: B +14.9CritPt: B +17.4AA-LCR: B +8.7Omniscience Acc.: B +4.6

Coding

TermBench: B +38.6SciCode: B +13.9WebDev Elo: B +180

Tool use & function calling

τ³-Banking: B +38.1

Arena

Arena Elo: B +41

Pick a different pair · Back to catalog