GLM-5.2 vs Qwen3.8-Max vs GPT-5.5
Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.
Cheapest blended
GLM-5.2
$2.15 / 1M tokens
Largest context
GPT-5.5
1.05M tokens
Fastest output
GLM-5.2
88 tokens/s
Top on SWE-bench Pro
Qwen3.8-Max
67.7%
Specifications
| Spec | GLM-5.2Z.ai | Qwen3.8-MaxAlibaba | GPT-5.5OpenAI |
|---|---|---|---|
| Provider | Z.ai | Alibaba | OpenAI |
| Released | 16 Jun 2026 | 2 Aug 2026 | 24 Apr 2026 |
| Licence | Open weight | Open weight | Proprietary |
| Context window | 1.05M tokens | 1M tokens | 1.05M tokens |
| Max output | 131K tokens | 131K tokens | 128K tokens |
| Input / 1M | $1.40* | $2.00* | $5.00* |
| Output / 1M | $4.40 | $6.00 | $30.00 |
| Cached input / 1M | $0.26 | $0.25 | $0.50 |
| Speed | 88 tokens/s | 39 tokens/s | 86 tokens/s |
| Input types | text | text, image, video | text, image |
| Best for | CodingAgentsLong context | CodingAgentsReasoningLong context | CodingAgents |
* GLM-5.2: Same price as GLM-5.3. Cached-input storage is free for a limited time.
* Qwen3.8-Max: International (Singapore) list price. Implicit cache $0.25; explicit cache creation $2.50, explicit cache read $0.17 (Qwen Cloud). Global deployment scope is cheaper at $1.65 / $4.951. A faster qwen3.8-max-prime mode costs about 2x. Snapshot qwen3.8-max-0902 added 2 Sep 2026.
* GPT-5.5: There is no separate cache-write charge (pre-GPT-5.6 models). Batch and Flex: $2.50 / $15. Fast mode: 2.5x, $12.50 / $75. Snapshot gpt-5.5-2026-04-23. GPT-5.5 Pro (gpt-5.5-pro) costs $30 / $180. Data residency +10%.
Benchmark scores
Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.
Price per 1M tokens
Context window and speed
Notes
- GLM-5.2
- Z.ai flagship for long-horizon tasks with a 1M-token context, 744B total and 40B active parameters. Open weights under plain MIT with no regional limits (commercial use allowed). Superseded by GLM-5.3 on 18 Aug 2026.
- Qwen3.8-Max
- Alibaba's current flagship, built on the open-weight Qwen3.8-2.4T-A95B (2.4T total, 95B active MoE). The API version adds vision input, non-thinking mode and 1M context by default. Weights license allows commercial use, but products over 100M MAU or $20M monthly revenue must show the model name, and Model-as-a-Service or AI coding/office assistant businesses with over $50M revenue in 12 months need a separate license from Qwen.
- GPT-5.5
- A frontier model for coding and complex professional work, the flagship before GPT-5.6.
Still torn between two?
Test them on your own prompts.
Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.