Qwen3.7-Max vs GLM-5.2 vs Qwen3.8-27B
Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.
Cheapest blended
Qwen3.8-27B
$1.13 / 1M tokens
Largest context
GLM-5.2
1.05M tokens
Fastest output
Qwen3.7-Max
203 tokens/s
Top on Terminal-Bench 2.1
GLM-5.2
81%
Specifications
| Spec | Qwen3.7-MaxAlibaba | GLM-5.2Z.ai | Qwen3.8-27BAlibaba |
|---|---|---|---|
| Provider | Alibaba | Z.ai | Alibaba |
| Released | 21 May 2026 | 16 Jun 2026 | 17 Aug 2026 |
| Licence | Proprietary | Open weight | Open weight |
| Context window | 1M tokens | 1.05M tokens | 1M tokens |
| Max output | 131K tokens | 131K tokens | 131K tokens |
| Input / 1M | $2.50* | $1.40* | $0.50* |
| Output / 1M | $7.50 | $4.40 | $3.00 |
| Cached input / 1M | $0.50 | $0.26 | $0.10 |
| Speed | 203 tokens/s | 88 tokens/s | 46 tokens/s |
| Input types | text | text | text, image, video |
| Best for | CodingAgentsReasoningLong context | CodingAgentsLong context | CodingAgentsVisionBudget |
* Qwen3.7-Max: International (Singapore) list price. Implicit cache $0.50; explicit cache read $0.25. Global deployment scope $1.65 / $4.951. Alibaba now lists Qwen3.7 under Legacy models; Qwen3.8-Max is cheaper ($2 / $6).
* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.
* Qwen3.8-27B: International (Singapore) price on Alibaba Model Studio / Qwen Cloud. Explicit cache read $0.05.
Benchmark scores
Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.
Price per 1M tokens
Context window and speed
Notes
- Qwen3.7-Max
- Largest model of the Qwen3.7 series, aimed at agent work, programming and long autonomous tasks. The qwen3.7-max alias points to the 2026-05-20 text-only snapshot; the 2026-06-08 snapshot adds image understanding. Not open weight.
- GLM-5.2
- Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.
- Qwen3.8-27B
- Dense 27B native vision-language model (images and video), open weights under Apache-2.0 with no extra commercial conditions. Weights support 262,144 tokens natively, extensible to 1M; the hosted API serves 1M by default.
Still torn between two?
Test them on your own prompts.
Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.