Skip to content
Free tool · Updated 2 Oct 2026

Qwen3.7-Max vs GLM-5.2 vs Qwen3.8-27B

Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.

Qwen3.7-MaxGLM-5.2Qwen3.8-27B

Cheapest blended

Qwen3.8-27B

$1.13 / 1M tokens

Largest context

GLM-5.2

1.05M tokens

Fastest output

Qwen3.7-Max

203 tokens/s

Top on Terminal-Bench 2.1

GLM-5.2

81%

Specifications

SpecQwen3.7-MaxAlibabaGLM-5.2Z.aiQwen3.8-27BAlibaba
ProviderAlibabaZ.aiAlibaba
Released21 May 202616 Jun 202617 Aug 2026
LicenceProprietaryOpen weightOpen weight
Context window1M tokens1.05M tokens1M tokens
Max output131K tokens131K tokens131K tokens
Input / 1M$2.50*$1.40*$0.50*
Output / 1M$7.50$4.40$3.00
Cached input / 1M$0.50$0.26$0.10
Speed203 tokens/s88 tokens/s46 tokens/s
Input typestexttexttext, image, video
Best forCodingAgentsReasoningLong contextCodingAgentsLong contextCodingAgentsVisionBudget

* Qwen3.7-Max: International (Singapore) list price. Implicit cache $0.50; explicit cache read $0.25. Global deployment scope $1.65 / $4.951. Alibaba now lists Qwen3.7 under Legacy models; Qwen3.8-Max is cheaper ($2 / $6).

* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.

* Qwen3.8-27B: International (Singapore) price on Alibaba Model Studio / Qwen Cloud. Explicit cache read $0.05.

Benchmark scores

Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.

Qwen3.7-MaxGLM-5.2Qwen3.8-27B
Terminal-Bench 2.174.5%81%73%
SWE-bench Pro60.6%62.1%61.7%
GPQA Diamond92.4%91.2%89.2%
Humanity’s Last Exam (no tools)41.4%40.5%30.8%

Price per 1M tokens

Qwen3.7-MaxGLM-5.2Qwen3.8-27B
Input$2.50$1.40$0.50
Output$7.50$4.40$3.00
Cached input$0.50$0.26$0.10

Context window and speed

Qwen3.7-MaxGLM-5.2Qwen3.8-27B
Context1M1.05M1M
Tokens/s2038846

Notes

Qwen3.7-Max
Largest model of the Qwen3.7 series, aimed at agent work, programming and long autonomous tasks. The qwen3.7-max alias points to the 2026-05-20 text-only snapshot; the 2026-06-08 snapshot adds image understanding. Not open weight.
GLM-5.2
Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.
Qwen3.8-27B
Dense 27B native vision-language model (images and video), open weights under Apache-2.0 with no extra commercial conditions. Weights support 262,144 tokens natively, extensible to 1M; the hosted API serves 1M by default.

Still torn between two?

Test them on your own prompts.

Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.