Skip to content
Free tool · Updated 2 Oct 2026

Qwen3.8-27B vs Qwen3.7-Max vs GLM-5.2

Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.

Qwen3.8-27BQwen3.7-MaxGLM-5.2

Cheapest blended

Qwen3.8-27B

$1.13 / 1M tokens

Largest context

GLM-5.2

1.05M tokens

Fastest output

Qwen3.7-Max

203 tokens/s

Top on Terminal-Bench 2.1

GLM-5.2

81%

Specifications

SpecQwen3.8-27BAlibabaQwen3.7-MaxAlibabaGLM-5.2Z.ai
ProviderAlibabaAlibabaZ.ai
Released17 Aug 202621 May 202616 Jun 2026
LicenceOpen weightProprietaryOpen weight
Context window1M tokens1M tokens1.05M tokens
Max output131K tokens131K tokens131K tokens
Input / 1M$0.50*$2.50*$1.40*
Output / 1M$3.00$7.50$4.40
Cached input / 1M$0.10$0.50$0.26
Speed46 tokens/s203 tokens/s88 tokens/s
Input typestext, image, videotexttext
Best forCodingAgentsVisionBudgetCodingAgentsReasoningLong contextCodingAgentsLong context

* Qwen3.8-27B: International (Singapore) price on Alibaba Model Studio / Qwen Cloud. Explicit cache read $0.05.

* Qwen3.7-Max: International (Singapore) list price. Implicit cache $0.50; explicit cache read $0.25. Global deployment scope $1.65 / $4.951. Alibaba now lists Qwen3.7 under Legacy models; Qwen3.8-Max is cheaper ($2 / $6).

* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.

Benchmark scores

Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.

Qwen3.8-27BQwen3.7-MaxGLM-5.2
Terminal-Bench 2.173%74.5%81%
SWE-bench Pro61.7%60.6%62.1%
GPQA Diamond89.2%92.4%91.2%
Humanity’s Last Exam (no tools)30.8%41.4%40.5%

Price per 1M tokens

Qwen3.8-27BQwen3.7-MaxGLM-5.2
Input$0.50$2.50$1.40
Output$3.00$7.50$4.40
Cached input$0.10$0.50$0.26

Context window and speed

Qwen3.8-27BQwen3.7-MaxGLM-5.2
Context1M1M1.05M
Tokens/s4620388

Notes

Qwen3.8-27B
Dense 27B native vision-language model (images and video), open weights under Apache-2.0 with no extra commercial conditions. Weights support 262,144 tokens natively, extensible to 1M; the hosted API serves 1M by default.
Qwen3.7-Max
Largest model of the Qwen3.7 series, aimed at agent work, programming and long autonomous tasks. The qwen3.7-max alias points to the 2026-05-20 text-only snapshot; the 2026-06-08 snapshot adds image understanding. Not open weight.
GLM-5.2
Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.

Still torn between two?

Test them on your own prompts.

Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.