GPT-6 Luna vs GLM-5.3 vs GPT-5.6 Luna
Pick two to four models to compare benchmark scores, prices, context windows and speed. The link updates as you pick, so you can share the exact comparison.
Cheapest blended
GPT-6 Luna
$0.20 / 1M tokens
Largest context
GPT-6 LunaGPT-5.6 Luna
Tie at 1.05M tokens
Fastest output
GPT-6 Luna
131 tokens/s
Top on DeepSWE v1.1
GPT-5.6 Luna
67.2%
Specifications
| Spec | GPT-6 LunaOpenAI | GLM-5.3Z.ai | GPT-5.6 LunaOpenAI |
|---|---|---|---|
| Provider | OpenAI | Z.ai | OpenAI |
| Released | 22 Sep 2026 | 18 Aug 2026 | 9 Jul 2026 |
| Licence | Proprietary | Open weight | Proprietary |
| Context window | 1.05M tokens | 1.05M tokens | 1.05M tokens |
| Max output | 128K tokens | 131K tokens | 128K tokens |
| Input / 1M | $0.10* | $1.40* | $0.20* |
| Output / 1M | $0.50 | $4.40 | $1.20 |
| Cached input / 1M | $0.01 | $0.26 | $0.02 |
| Speed | 131 tokens/s | 70 tokens/s | 124 tokens/s |
| Input types | text, image | text | text, image |
| Best for | BudgetAgents | CodingAgentsLong context | Budget |
* GPT-6 Luna: Batch and Flex: $0.05 / $0.25. Fast mode: $0.20 / $1.00. Cache writes 1.25x input; 30-minute minimum cache life. Data residency +10%. Launched at 50% below GPT-5.6 Luna's price.
* GLM-5.3: Cached-input storage is free for a limited time.
* GPT-5.6 Luna: Cut 80% on 30 Jul 2026 (launch price $1 in / $6 out). Batch and Flex: $0.10 / $0.60. Fast mode: $0.40 / $2.40. Cache writes 1.25x input; 30-minute minimum cache life. Data residency +10%.
Benchmark scores
Officially reported scores only. A model with no published score on a benchmark shows a line, not a zero.
Price per 1M tokens
Context window and speed
Notes
- GPT-6 Luna
- OpenAI's most efficient model, for focused, high-volume tasks.
- GLM-5.3
- Z.ai's current flagship: same 744B-total / 40B-active base as GLM-5.2, improved only by post-training, with large gains in coding and cyber-security tasks. Weights license is MIT-style and allows commercial use; only a Model-as-a-Service operator with over $10B revenue in 12 months must pass a Z.ai security review first.
- GPT-5.6 Luna
- The fastest and most affordable GPT-5.6 tier, for cost-sensitive, high-volume workloads. It is the successor to the earlier 'nano' tier.
Still torn between two?
Test them on your own prompts.
Benchmarks are someone else’s work. On a free call we look at yours and agree how to check quality, cost and latency for the models you are weighing.