Skip to content

GPT-5.4

OpenAI · Released 5 Mar 2026 · gpt-5.4

Compare this model
ProprietaryCodingAgents

$2.50

Input price / 1M tokens

#45 of 61, cheapest first

$15.00

Output price / 1M tokens

#46 of 61, cheapest first

1.05M

Context window

#2 of 71

88

Output tokens / second

#34 of 61

75.1%

Terminal-Bench 2.0

#2 of 7

57.7%

SWE-bench Pro

#13 of 16

75%

OSWorld-Verified

#6 of 8

92.8%

GPQA Diamond

#7 of 27

Summary

GPT-5.4 is a proprietary model from OpenAI, released on 5 Mar 2026.

At $2.50 input and $15.00 output per million tokens, it is #45 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #2 of 71.

Measured output speed is 88 tokens per second, #34 of 61.

A frontier model for professional work and coding, and OpenAI's first general-purpose model with native computer use. It is now a lower-cost alternative to GPT-5.5.

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.0xhigh effort75.1%#2 of 7OpenAI GPT-5.5 launch post (comparison column)Self-reported
SWE-bench Proxhigh effort57.7%#13 of 16OpenAI GPT-5.5 launch post (comparison column)Self-reported
OSWorld-Verifiedxhigh effort75%#6 of 8OpenAI GPT-5.5 launch post (comparison column)Self-reported
GPQA Diamondxhigh effort92.8%#7 of 27OpenAI GPT-5.5 launch post (comparison column)Self-reported
BrowseCompxhigh effort82.7%#7 of 7OpenAI GPT-5.5 launch post (comparison column)Self-reported
ARC-AGI-2xhigh effort73.3%#4 of 6OpenAI GPT-5.5 launch post (comparison column)Self-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with GPT-5.4, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.0

#2 of 7

Terminal tasks, 2025 set

SWE-bench Pro

#13 of 16

Harder multi-file engineering tasks

OSWorld-Verified

#6 of 8

Using a desktop computer

GPQA Diamond

#7 of 27

Graduate-level science questions

Compare GPT-5.4 side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on GPT-5.4?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.