By the toolist.ai editors. Researched from public sources, not hands-on tested by toolist.ai. Updated October 6, 2026; prices checked October 6, 2026.

GPT-6 Astra: price, context window and benchmarks

GPT-6 Astra is OpenAI's most capable GPT-6 model, at $10 per million input tokens and $50 per million output tokens, five times GPT-6.1 Sol. OpenAI aims it at the hardest reasoning, coding and research work; on Artificial Analysis it scores 52.7 on the Intelligence Index, a point above GPT-6.1 Sol.

Best for

Key facts

FactGPT-6 AstraSource
ProviderOpenAIOpenAI API docs: Pricing
API model namegpt-6-astraOpenAI API docs: GPT-6 Astra
ReleasedSeptember 3, 2026OpenAI API docs: Changelog
Input price$10 per million tokensOpenAI API docs: Pricing
Cached input price$1 per million tokensOpenAI API docs: Pricing
Output price$50 per million tokensOpenAI API docs: Pricing
Context window1,050,000 tokensOpenAI API docs: GPT-6 Astra
Maximum output128,000 tokensOpenAI API docs: GPT-6 Astra
Knowledge cutoffApril 30, 2026OpenAI API docs: GPT-6 Astra
Input and outputText and images in; text outOpenAI API docs: GPT-6 Astra
ReasoningAlways reasons. Effort levels low to max; "none" is not supportedOpenAI API docs: Reasoning models
API and cloud platformsOpenAI API (Responses, Chat Completions and Batch)OpenAI API docs: GPT-6 Astra
In the appsChatGPT Plus, Pro, Business, Enterprise and Edu, in ChatGPT Work and Codex; not on Free or GoChatGPT docs: Pricing and plan access

Prompts over 272,000 input tokens are billed at twice the input rate and 1.5 times the output rate for the whole request. Batch and Flex cost half; Fast mode costs double and Ultrafast six times.

What GPT-6 Astra costs a month

Three typical workloads, priced at the standard API rates above. We calculated these; the provider does not publish them.

WorkloadTokens a monthCost a month
Support chatbot
100,000 replies, each 3,000 input and 400 output tokens, no caching
300M in, 40M out $5,000
Document Q&A
5,000 questions over the same 60,000-token document set (90% cache hits), 800 output tokens each
300M in, 4M out $770
Coding agent
200 tasks, each 1.5 million input tokens (90% cache hits) and 50,000 output tokens
300M in, 10M out $1,070

Reasoning models bill their thinking as output tokens, so heavy reasoning can raise the output count several times. Cache-write charges are not included.

Benchmarks

BenchmarkScoreSource
Artificial Analysis Intelligence Index
max effort
52.7 Artificial Analysis: GPT-6 Astra (max)
Intelligence Index at everyday effort
medium effort; OpenAI publishes no default
49.6 Artificial Analysis: GPT-6 Astra (medium)
Coding Agent Index
Codex harness, max effort
61.6 Artificial Analysis: Coding agents
Terminal-Bench 4.0
max effort
59.1% Artificial Analysis: GPT-6.1 Sol vs GPT-6 Astra
Humanity's Last Exam
max effort
54.7% Artificial Analysis: GPT-6.1 Sol vs GPT-6 Astra
GDPval-AA v2.1 (Elo)
max effort
1542 Artificial Analysis: GPT-6.1 Sol vs GPT-6 Astra
SciCode
max effort
56.5% Artificial Analysis: GPT-6.1 Sol vs GPT-6 Astra
Terminal-Bench-Science 0.1
max effort
63.3% Artificial Analysis: GPT-6 Astra (max)
GPQA Diamond
max effort
96.1% Artificial Analysis: GPT-6 Astra (max)

Speed: about 50.9 output tokens a second, 5.5 seconds to the first token (medium effort, OpenAI API; Artificial Analysis: GPT-6 Astra (medium)).

Benchmark and speed data marked as such come from Artificial Analysis.

What changed

The first GPT-6 model and a new top tier above Sol. OpenAI's guide describes Astra as its highest-intelligence model, Sol as balanced between speed, cost and intelligence, and Luna as the fastest and cheapest. (OpenAI API docs: Using GPT-6)

GPT-6 Astra compared

Other models

All prices side by side: AI model API price table. How we research: our method.

Sources