Skip to content

GPT-6 Astra

OpenAI · Released 3 Sep 2026 · gpt-6-astra

Compare this model
ProprietaryCodingAgentsReasoningResearchLong context

$10.00

Input price / 1M tokens

#59 of 61, cheapest first

$50.00

Output price / 1M tokens

#59 of 61, cheapest first

1.05M

Context window

#2 of 71

51

Output tokens / second

#52 of 61

57.9%

Terminal-Bench 4.0

#4 of 10

74.1%

DeepSWE v1.1

#4 of 17

72.6%

OSWorld 2.0

#3 of 10

91.5%

BrowseComp

#1 of 7

Summary

GPT-6 Astra is a proprietary model from OpenAI, released on 3 Sep 2026.

At $10.00 input and $50.00 output per million tokens, it is #59 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #2 of 71.

Measured output speed is 51 tokens per second, #52 of 61.

OpenAI's most capable model, for the hardest end-to-end work: complex reasoning, coding, computer use, research and document creation.

Benchmarks · 12 reported

BenchmarkScoreBarRankSource
Terminal-Bench 4.0max score at any effort57.9%#4 of 10OpenAI launch postSelf-reported
DeepSWE v1.1max score at any effort74.1%#4 of 17OpenAI launch postSelf-reported
OSWorld 2.0v2026.08.08, offline set, partial score72.6%#3 of 10OpenAI launch postSelf-reported
Agents' Last Exammax score at any effort59.3%–OpenAI launch postSelf-reported
BrowseCompmax score at any effort91.5%#1 of 7OpenAI launch postSelf-reported
AutomationBenchmax score at any effort41.4%#1 of 8OpenAI launch postSelf-reported
GPQA Diamondmax score at any effort96%#1 of 27OpenAI launch postSelf-reported
Humanity’s Last Exam (with tools)with tools57.2%#9 of 16OpenAI launch postSelf-reported
FrontierMath Tier 4 (v2)max score at any effort97.6%–OpenAI launch postSelf-reported
ARC-AGI-2max score at any effort95%#1 of 6OpenAI launch postSelf-reported
ARC-AGI-3run with OpenAI's Responses API harness (two settings changed, per footnote 1)99.9%–OpenAI launch postSelf-reported
ScreenSpot-Prono tools92.7%–OpenAI launch postSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with GPT-6 Astra, and the closest scores.

Compare 4 side by side

Terminal-Bench 4.0

#4 of 10

Long jobs in a real terminal

DeepSWE v1.1

#4 of 17

Long coding tasks in real repos

OSWorld 2.0

#3 of 10

Long desktop and web workflows

BrowseComp

#1 of 7

Finding hard facts on the web

Compare GPT-6 Astra side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on GPT-6 Astra?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.