Skip to content

GLM-5.3

Z.ai · Released 18 Aug 2026 · glm-5.3

Compare this model
Open weightCodingAgentsLong context

$1.40

Input price / 1M tokens

#32 of 61, cheapest first

$4.40

Output price / 1M tokens

#27 of 61, cheapest first

1.05M

Context window

#11 of 71

70

Output tokens / second

#45 of 61

88.2%

Terminal-Bench 2.1

#5 of 17

66.9%

DeepSWE v1.1

#12 of 17

62.5%

Humanity’s Last Exam (with tools)

#7 of 16

Summary

GLM-5.3 is an open-weight model from Z.ai, released on 18 Aug 2026.

At $1.40 input and $4.40 output per million tokens, it is #32 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #11 of 71.

Measured output speed is 70 tokens per second, #45 of 61.

Z.ai's current flagship: same 744B-total / 40B-active base as GLM-5.2, improved only by post-training, with large gains in coding and cyber-security tasks. Weights license is MIT-style and allows commercial use; only a Model-as-a-Service operator with over $10B revenue in 12 months must pass a Z.ai security review first.

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.1Claude Code harness, max effort88.2%#5 of 17GLM-5.3 model cardSelf-reported
Terminal-Bench 3.0Claude Code harness, max effort, avg@328.3%–GLM-5.3 model cardSelf-reported
DeepSWE v1.1mini-swe-agent66.9%#12 of 17GLM-5.3 model cardSelf-reported
CyberGymClaude Code harness, max effort, Pass@184.5%–GLM-5.3 model cardSelf-reported
Humanity’s Last Exam (with tools)with tools62.5%#7 of 16GLM-5.3 model cardSelf-reported
Toolathlon VerifiedPass@1, avg of 3 runs73%–GLM-5.3 model cardSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with GLM-5.3, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#5 of 17

Terminal tasks, earlier set

DeepSWE v1.1

#12 of 17

Long coding tasks in real repos

Humanity’s Last Exam (with tools)

#7 of 16

Expert questions, with search and code

Compare GLM-5.3 side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on GLM-5.3?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.