Qwen3.8-27B vs Qwen3.7-Max vs GLM-5.2
Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.
Cheapest blended
Qwen3.8-27B
$1.13 / 1M tokens
Largest context
GLM-5.2
1.05M tokens
Fastest output
Qwen3.7-Max
203 tokens/s
Top on Terminal-Bench 2.1
GLM-5.2
81%
Specifications
| Spec | Qwen3.8-27BAlibaba | Qwen3.7-MaxAlibaba | GLM-5.2Z.ai |
|---|---|---|---|
| Provider | Alibaba | Alibaba | Z.ai |
| Released | 17 Aug 2026 | 21 May 2026 | 16 Jun 2026 |
| Licence | Open weight | Proprietary | Open weight |
| Context window | 1M tokens | 1M tokens | 1.05M tokens |
| Max output | 131K tokens | 131K tokens | 131K tokens |
| Input / 1M | $0.50* | $2.50* | $1.40* |
| Output / 1M | $3.00 | $7.50 | $4.40 |
| Cached input / 1M | $0.10 | $0.50 | $0.26 |
| Speed | 46 tokens/s | 203 tokens/s | 88 tokens/s |
| Input types | text, image, video | text | text |
| Best for | CodingAgentsVisionBudget | CodingAgentsReasoningLong context | CodingAgentsLong context |
* Qwen3.8-27B: International (Singapore) price on Alibaba Model Studio / Qwen Cloud. Explicit cache read $0.05.
* Qwen3.7-Max: International (Singapore) list price. Implicit cache $0.50; explicit cache read $0.25. Global deployment scope $1.65 / $4.951. Alibaba now lists Qwen3.7 under Legacy models; Qwen3.8-Max is cheaper ($2 / $6).
* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.
Benchmark scores
Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.
Price per 1M tokens
Context window and speed
Notes
- Qwen3.8-27B
- Dense 27B native vision-language model (images and video), open weights under Apache-2.0 with no extra commercial conditions. Weights support 262,144 tokens natively, extensible to 1M; the hosted API serves 1M by default.
- Qwen3.7-Max
- Largest model of the Qwen3.7 series, aimed at agent work, programming and long autonomous tasks. The qwen3.7-max alias points to the 2026-05-20 text-only snapshot; the 2026-06-08 snapshot adds image understanding. Not open weight.
- GLM-5.2
- Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.
Still torn between two?
Test them on your own prompts.
Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.