Skip to content

DeepSeek V4.1 Flash

DeepSeek · Released 10 Sep 2026 · deepseek-flash

Compare this model
Open weightCodingAgentsBudgetVisionLong context

$0.30

Input price / 1M tokens

#9 of 61, cheapest first

$1.20

Output price / 1M tokens

#8 of 61, cheapest first

1.05M

Context window

#11 of 71

209

Output tokens / second

#8 of 61

90.6%

Terminal-Bench 2.1

#1 of 17

74.2%

DeepSWE v1.1

#3 of 17

90.9%

GPQA Diamond

#14 of 27

36.8%

Humanity’s Last Exam (no tools)

#12 of 23

Summary

DeepSeek V4.1 Flash is an open-weight model from DeepSeek, released on 10 Sep 2026.

At $0.30 input and $1.20 output per million tokens, it is #9 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #11 of 71.

Measured output speed is 209 tokens per second, #8 of 61.

Smallest model in DeepSeek's new architecture family: a 552B-parameter MoE (Causal Encoder-Decoder) that activates 8B parameters per token for input and 16B for output, with native image understanding. DeepSeek says it beats V4-Pro on its benchmarks. Open weights under MIT (commercial use allowed).

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.1DeepSeek Harness minimal mode, reasoning_effort=10090.6%#1 of 17DeepSeek-V4.1-Flash model cardSelf-reported
DeepSWE v1.1mini-SWE harness, max effort74.2%#3 of 17DeepSeek-V4.1-Flash model cardSelf-reported
GPQA DiamondPass@1, max effort90.9%#14 of 27DeepSeek-V4.1-Flash model cardSelf-reported
Humanity’s Last Exam (no tools)no tools, full set (39.1 on text-only subset)36.8%#12 of 23DeepSeek-V4.1-Flash model cardSelf-reported
Humanity’s Last Exam (with tools)with tools63.9%#4 of 16DeepSeek-V4.1-Flash model cardSelf-reported
Codeforcesmax effort3,471Elo-style–DeepSeek-V4.1-Flash model cardSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with DeepSeek V4.1 Flash, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#1 of 17

Terminal tasks, earlier set

DeepSWE v1.1

#3 of 17

Long coding tasks in real repos

GPQA Diamond

#14 of 27

Graduate-level science questions

Humanity’s Last Exam (no tools)

#12 of 23

Expert questions, no search or code

Compare DeepSeek V4.1 Flash side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on DeepSeek V4.1 Flash?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.