Skip to content

Kimi K3

Moonshot · Released Jul 2026 · kimi-k3

Compare this model
Open weightCodingAgentsReasoningLong contextMultimodal

$3.00

Input price / 1M tokens

#47 of 61, cheapest first

$15.00

Output price / 1M tokens

#46 of 61, cheapest first

1.05M

Context window

#11 of 71

34

Output tokens / second

#61 of 61

88.3%

Terminal-Bench 2.1

#4 of 17

67.5%

DeepSWE v1.1

#10 of 17

93.5%

GPQA Diamond

#5 of 27

43.5%

Humanity’s Last Exam (no tools)

#6 of 23

Summary

Kimi K3 is an open-weight model from Moonshot, released on Jul 2026.

At $3.00 input and $15.00 output per million tokens, it is #47 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #11 of 71.

Measured output speed is 34 tokens per second, #61 of 61.

Moonshot's flagship and the first open 3T-class model: 2.8T total, 104B active parameters, native vision, 1M context. Kimi K3 License allows commercial use, but Model-as-a-Service operators with over $20M revenue in 12 months need a separate agreement, and products over 100M MAU or $20M monthly revenue must display "Kimi K3".

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.1max effort88.3%#4 of 17Kimi-K3 model cardSelf-reported
DeepSWE v1.1max effort67.5%#10 of 17Kimi-K3 model cardSelf-reported
GPQA Diamondmax effort93.5%#5 of 27Kimi-K3 model cardSelf-reported
Humanity’s Last Exam (no tools)full set, no tools, max effort43.5%#6 of 23Kimi-K3 model cardSelf-reported
Humanity’s Last Exam (with tools)full set, with tools, max effort56%#11 of 16Kimi-K3 model cardSelf-reported
OSWorld-Verifiedmax effort84.8%#1 of 8Kimi-K3 model cardSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with Kimi K3, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#4 of 17

Terminal tasks, earlier set

DeepSWE v1.1

#10 of 17

Long coding tasks in real repos

GPQA Diamond

#5 of 27

Graduate-level science questions

Humanity’s Last Exam (no tools)

#6 of 23

Expert questions, no search or code

Compare Kimi K3 side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on Kimi K3?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.