Skip to content

Qwen3.8-Max

Alibaba · Released 2 Aug 2026 · qwen3.8-max

Compare this model
Open weightCodingAgentsReasoningLong contextMultimodal

$2.00

Input price / 1M tokens

#36 of 61, cheapest first

$6.00

Output price / 1M tokens

#32 of 61, cheapest first

1M

Context window

#29 of 71

39

Output tokens / second

#57 of 61

86.6%

Terminal-Bench 2.1

#8 of 17

67.7%

SWE-bench Pro

#1 of 16

92.6%

GPQA Diamond

#8 of 27

43.6%

Humanity’s Last Exam (no tools)

#5 of 23

Summary

Qwen3.8-Max is an open-weight model from Alibaba, released on 2 Aug 2026.

At $2.00 input and $6.00 output per million tokens, it is #36 of 61 on input price, cheapest first.

Its 1M-token context window ranks #29 of 71.

Measured output speed is 39 tokens per second, #57 of 61.

Alibaba's current flagship, built on the open-weight Qwen3.8-2.4T-A95B (2.4T total, 95B active MoE). The API version adds vision input, non-thinking mode and 1M context by default. Weights license allows commercial use, but products over 100M MAU or $20M monthly revenue must show the model name, and Model-as-a-Service or AI coding/office assistant businesses with over $50M revenue in 12 months need a separate license from Qwen.

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.186.6%#8 of 17Qwen3.8-2.4T-A95B model cardSelf-reported
SWE-bench Pro67.7%#1 of 16Qwen3.8-2.4T-A95B model cardSelf-reported
GPQA Diamond92.6%#8 of 27Qwen3.8-2.4T-A95B model cardSelf-reported
Humanity’s Last Exam (no tools)no tools43.6%#5 of 23Qwen3.8-2.4T-A95B model cardSelf-reported
Humanity’s Last Exam (with tools)with tools56.2%#10 of 16Qwen3.8-2.4T-A95B model cardSelf-reported
Toolathlon VerifiedPass@172.5%–Qwen3.8-2.4T-A95B model cardSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with Qwen3.8-Max, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#8 of 17

Terminal tasks, earlier set

SWE-bench Pro

#1 of 16

Harder multi-file engineering tasks

GPQA Diamond

#8 of 27

Graduate-level science questions

Humanity’s Last Exam (no tools)

#5 of 23

Expert questions, no search or code

Compare Qwen3.8-Max side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on Qwen3.8-Max?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.