Skip to content

GPT-5.6 Sol

OpenAI · Released 9 Jul 2026 · gpt-5.6-sol

Compare this model
ProprietaryCodingAgentsReasoning

$4.00

Input price / 1M tokens

#50 of 61, cheapest first

$20.00

Output price / 1M tokens

#50 of 61, cheapest first

1.05M

Context window

#2 of 71

79

Output tokens / second

#41 of 61

64.6%

SWE-bench Pro

#2 of 16

72.7%

DeepSWE v1.1

#6 of 17

88.8%

Terminal-Bench 2.1

#3 of 17

62.6%

OSWorld 2.0

#5 of 10

Summary

GPT-5.6 Sol is a proprietary model from OpenAI, released on 9 Jul 2026.

At $4.00 input and $20.00 output per million tokens, it is #50 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #2 of 71.

Measured output speed is 79 tokens per second, #41 of 61.

Flagship of the GPT-5.6 family, for complex professional work, coding and long-horizon agentic tasks. It adds max reasoning effort and Pro mode.

Benchmarks · 8 reported

BenchmarkScoreBarRankSource
SWE-bench Pro64.6%#2 of 16OpenAI launch postSelf-reported
DeepSWE v1.172.7%#6 of 17OpenAI launch postSelf-reported
Terminal-Bench 2.1single agent (91.9% with ultra multi-agent)88.8%#3 of 17OpenAI launch postSelf-reported
OSWorld 2.062.6%#5 of 10OpenAI launch postSelf-reported
BrowseCompsingle agent (92.2% with ultra multi-agent)90.4%#2 of 7OpenAI launch postSelf-reported
GPQA Diamond94.6%#2 of 27OpenAI launch postSelf-reported
FrontierMath Tier 4 (v2)83%–OpenAI launch postSelf-reported
Agents' Last Examresults table value52.7%–OpenAI launch postSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with GPT-5.6 Sol, and the closest scores.

Compare 4 side by side

SWE-bench Pro

#2 of 16

Harder multi-file engineering tasks

DeepSWE v1.1

#6 of 17

Long coding tasks in real repos

Terminal-Bench 2.1

#3 of 17

Terminal tasks, earlier set

OSWorld 2.0

#5 of 10

Long desktop and web workflows

Compare GPT-5.6 Sol side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on GPT-5.6 Sol?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.