Skip to content

Qwen3.8-Flash

Alibaba · Released 26 Aug 2026 · qwen3.8-flash

Compare this model
Open weightBudgetCodingAgentsMultimodalLong context

$0.15

Input price / 1M tokens

#3 of 61, cheapest first

$0.47

Output price / 1M tokens

#1 of 61, cheapest first

1M

Context window

#29 of 71

–

Output tokens / second

62.5%

SWE-bench Pro

#5 of 16

58.7%

DeepSWE v1.1

#16 of 17

91.7%

GPQA Diamond

#12 of 27

35.9%

Humanity’s Last Exam (no tools)

#13 of 23

Summary

Qwen3.8-Flash is an open-weight model from Alibaba, released on 26 Aug 2026.

At $0.15 input and $0.47 output per million tokens, it is #3 of 61 on input price, cheapest first.

Its 1M-token context window ranks #29 of 71.

Fast multimodal Qwen model based on the open-weight Qwen3.8-Flash-Next (125B total, 6B active, plus 51B n-gram embedding), an experimental preview of the Qwen4 architecture. The Qwen Community License allows commercial use with a name-display rule above 100M MAU or $20M monthly revenue, but any Model-as-a-Service or AI coding/office assistant business needs a separate license from Qwen regardless of size.

Benchmarks · 5 reported

BenchmarkScoreBarRankSource
SWE-bench ProQwen3.8-Flash-Next open weights62.5%#5 of 16Qwen3.8-Flash-Next model cardSelf-reported
DeepSWE v1.1Qwen3.8-Flash-Next open weights58.7%#16 of 17Qwen3.8-Flash-Next model cardSelf-reported
GPQA DiamondQwen3.8-Flash-Next open weights91.7%#12 of 27Qwen3.8-Flash-Next model cardSelf-reported
Humanity’s Last Exam (no tools)no tools, Qwen3.8-Flash-Next open weights35.9%#13 of 23Qwen3.8-Flash-Next model cardSelf-reported
Toolathlon VerifiedPass@1, Qwen3.8-Flash-Next open weights73.5%–Qwen3.8-Flash-Next model cardSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with Qwen3.8-Flash, and the closest scores.

Compare 4 side by side

SWE-bench Pro

#5 of 16

Harder multi-file engineering tasks

DeepSWE v1.1

#16 of 17

Long coding tasks in real repos

GPQA Diamond

#12 of 27

Graduate-level science questions

Humanity’s Last Exam (no tools)

#13 of 23

Expert questions, no search or code

Compare Qwen3.8-Flash side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on Qwen3.8-Flash?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.