Skip to content

Qwen3.7-Max

Alibaba · Released 21 May 2026 · qwen3.7-max

Compare this model
ProprietaryCodingAgentsReasoningLong context

$2.50

Input price / 1M tokens

#45 of 61, cheapest first

$7.50

Output price / 1M tokens

#35 of 61, cheapest first

1M

Context window

#29 of 71

203

Output tokens / second

#10 of 61

74.5%

Terminal-Bench 2.1

#14 of 17

60.6%

SWE-bench Pro

#8 of 16

92.4%

GPQA Diamond

#9 of 27

41.4%

Humanity’s Last Exam (no tools)

#8 of 23

Summary

Qwen3.7-Max is a proprietary model from Alibaba, released on 21 May 2026.

At $2.50 input and $7.50 output per million tokens, it is #45 of 61 on input price, cheapest first.

Its 1M-token context window ranks #29 of 71.

Measured output speed is 203 tokens per second, #10 of 61.

Largest model of the Qwen3.7 series, aimed at agent work, programming and long autonomous tasks. The qwen3.7-max alias points to the 2026-05-20 text-only snapshot; the 2026-06-08 snapshot adds image understanding. Not open weight.

Benchmarks · 4 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.174.5%#14 of 17Qwen3.8-2.4T-A95B model card (Qwen3.7-Max column)Self-reported
SWE-bench Pro60.6%#8 of 16Qwen3.8-2.4T-A95B model card (Qwen3.7-Max column)Self-reported
GPQA Diamond92.4%#9 of 27Qwen3.8-2.4T-A95B model card (Qwen3.7-Max column)Self-reported
Humanity’s Last Exam (no tools)no tools41.4%#8 of 23Qwen3.8-2.4T-A95B model card (Qwen3.7-Max column)Self-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with Qwen3.7-Max, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#14 of 17

Terminal tasks, earlier set

SWE-bench Pro

#8 of 16

Harder multi-file engineering tasks

GPQA Diamond

#9 of 27

Graduate-level science questions

Humanity’s Last Exam (no tools)

#8 of 23

Expert questions, no search or code

Compare Qwen3.7-Max side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on Qwen3.7-Max?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.