Skip to content
Free tool · Updated 2 Oct 2026

GLM-5.2 vs Qwen3.8-Max vs GPT-5.5

Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.

GLM-5.2Qwen3.8-MaxGPT-5.5

Cheapest blended

GLM-5.2

$2.15 / 1M tokens

Largest context

GPT-5.5

1.05M tokens

Fastest output

GLM-5.2

88 tokens/s

Top on SWE-bench Pro

Qwen3.8-Max

67.7%

Specifications

SpecGLM-5.2Z.aiQwen3.8-MaxAlibabaGPT-5.5OpenAI
ProviderZ.aiAlibabaOpenAI
Released16 Jun 20262 Aug 202624 Apr 2026
LicenceOpen weightOpen weightProprietary
Context window1.05M tokens1M tokens1.05M tokens
Max output131K tokens131K tokens128K tokens
Input / 1M$1.40*$2.00*$5.00*
Output / 1M$4.40$6.00$30.00
Cached input / 1M$0.26$0.25$0.50
Speed88 tokens/s39 tokens/s86 tokens/s
Input typestexttext, image, videotext, image
Best forCodingAgentsLong contextCodingAgentsReasoningLong contextCodingAgents

* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.

* Qwen3.8-Max: International (Singapore) list price. Implicit cache $0.25; explicit cache creation $2.50, explicit cache read $0.17 (Qwen Cloud). Global deployment scope is cheaper at $1.65 / $4.951. A faster qwen3.8-max-prime mode costs about 2x. Snapshot qwen3.8-max-0902 added 2 Sep 2026.

* GPT-5.5: There is no separate cache-write charge (pre-GPT-5.6 models). Batch and Flex: $2.50 / $15. Fast mode: 2.5x, $12.50 / $75. Snapshot gpt-5.5-2026-04-23. GPT-5.5 Pro (gpt-5.5-pro) costs $30 / $180. Data residency +10%.

Benchmark scores

Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.

GLM-5.2Qwen3.8-MaxGPT-5.5
SWE-bench Pro62.1%67.7%58.6%
Terminal-Bench 2.181%86.6% no published score
GPQA Diamond91.2%92.6%93.6%
Humanity’s Last Exam (no tools)40.5%43.6%41.4%
Humanity’s Last Exam (with tools)54.7%56.2%52.2%

Price per 1M tokens

GLM-5.2Qwen3.8-MaxGPT-5.5
Input$1.40$2.00$5.00
Output$4.40$6.00$30.00
Cached input$0.26$0.25$0.50

Context window and speed

GLM-5.2Qwen3.8-MaxGPT-5.5
Context1.05M1M1.05M
Tokens/s883986

Notes

GLM-5.2
Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.
Qwen3.8-Max
Alibaba's current flagship, built on the open-weight Qwen3.8-2.4T-A95B (2.4T total, 95B active MoE). The API version adds vision input, non-thinking mode and 1M context by default. Weights license allows commercial use, but products over 100M MAU or $20M monthly revenue must show the model name, and Model-as-a-Service or AI coding/office assistant businesses with over $50M revenue in 12 months need a separate license from Qwen.
GPT-5.5
A frontier model for coding and complex professional work, the flagship before GPT-5.6.

Still torn between two?

Test them on your own prompts.

Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.