DataLearner logo
GP

GPT-6 Astra

PreviewReasoning modelCoding modelGPT

OpenAI GPT-6 Astra

Also known as: GPT-6 / GPT Astra / OpenAI Astra / Astra

Release date: 2026-09-03Updated: 2026-09-04Knowledge cutoff: 2026-04-30Views: 433
Live demoGitHubHugging FaceCompare
Parameters
No data
Context length
1.05M
Multilingual
Supported
Reasoning ability
5/5

OpenAI released GPT-6 Astra on September 3, 2026 as its flagship model for complex reasoning, coding, computer use, research, and document work. The official API ID is gpt-6-astra, with a 1.05M-token context window, 128K maximum output, text and image input, and low, medium, high, xhigh, and max reasoning effort levels.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-6 Astra

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · Medium (Default)Thinking Level · LowThinking Level · HighThinking Level · Extra-HighThinking Level · Max
Context length
1.05M tokens
Max output length
128K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2026-09-03
Model file size
No data
MoE architecture
No
Total params / Active params
No data / Not applicable
Knowledge cutoff
2026-04-30
GPT-6 Astra

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
N/A
Hugging Face
N/A
Live demo
N/A
GPT-6 Astra

Official resources

DataLearnerAI blog
N/A
GPT-6 Astra

API details

API speed
4/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
TextContext <= 272000$10.00/ 1M$50.00/ 1M
TextContext > 272000$20.00/ 1M$75.00/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$12.50/ 1M
Context <= 272000
$1.00/ 1M
Context <= 272000
Text-$25.00/ 1M
Context > 272000
$2.00/ 1M
Context > 272000

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

GPT-6 Astra

Benchmark Results

GPT-6 Astra currently shows benchmark results led by GPQA Diamond (1 / 271, score 96), ARC-AGI-1 (1 / 91, score 98.50), ARC-AGI-2 (1 / 85, score 95). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

8 evaluations
Benchmark / mode
Score
Rank/total
98.50
1 / 91
95
1 / 85
62.70
1 / 16
61.20
7 / 28
HLE
MaxTools
57.20
12 / 190
2.40
1 / 1
0
1 / 1

General Evaluation

5 evaluations
Benchmark / mode
Score
Rank/total

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
MaxTools
91.50
1 / 56

Coding and Software Engineer

7 evaluations
Benchmark / mode
Score
Rank/total
85.40
1 / 1
DeepSWE
MaxTools
74.10
2 / 35
67
4 / 4
64.50
1 / 6
63.90
1 / 1
53.30
1 / 1

AI Agent - Tool Usage

10 evaluations
Benchmark / mode
Score
Rank/total
ExploitBench
MaxTools
100
1 / 2
100
1 / 1
OSWorld 2.0
MaxTools
72.60
2 / 9
64.60
1 / 11
57.90
1 / 13
42.40
1 / 1
41.40
5 / 14
0
1 / 1

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
59.30
1 / 15

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total

Productivity Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
BenchCAD
MaxTools
95.90
1 / 1
50
1 / 1
40.90
1 / 1

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total

Truthfulness Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
4.20
1 / 1

Long Context

2 evaluations
Benchmark / mode
Score
Rank/total

Compare with other models

GPT-6 Astra

Publisher

OpenAI GPT-6 Astra

Model Overview

GPT-6 is officially GPT-6 Astra

GPT-6 Astra is OpenAI's flagship reasoning model released on September 3, 2026. Its official API model ID is gpt-6-astra. OpenAI positions it for complex reasoning, software engineering, computer use, research, and document creation, with an emphasis on carrying multistep work from an initial request to a finished result. The former GPT-6 and Astra rumor entries in this catalog now resolve to this single official product identity.

Specifications and modalities

  • Context window: 1,050,000 tokens
  • Maximum output: 128,000 tokens
  • Knowledge cutoff: April 30, 2026
  • Input: text and images
  • Output: text
  • Reasoning effort: low, medium, high, xhigh, and max; none is not supported

OpenAI has not disclosed total parameters, active parameters, or architecture details, so those catalog fields remain empty. The documentation confirms multilingual capability but does not provide a verified language count or complete BCP-47 list.

Agent and tool capabilities

GPT-6 Astra supports streaming, function calling, Structured Outputs, and Responses API tools including web search, file search, image generation, Code Interpreter, hosted shell, Apply Patch, Skills, computer use, MCP, and tool search. New platform capabilities include asynchronous tool calling, mid-turn steering over WebSockets, and configuration_update items that change reasoning effort while preserving the cached prompt prefix. Chat Completions is supported, but tool calling requires the Responses API.

Pricing

Standard short-context rates per million tokens are $10 input, $1 cached input, $12.50 cache writes, and $50 output. When a request contains more than 272K input tokens, the entire request uses long-context rates of $20 input, $2 cached input, $25 cache writes, and $75 output. Batch and Flex cost 50% of Standard rates, while Fast mode costs twice the applicable rate. Eligible regional processing may add 10%.

Availability

GPT-6 Astra shipped on September 3, 2026 as a staged rollout. Access went first to approved customers in Daybreak, OpenAI's application-only cybersecurity program, with ChatGPT Plus, Pro, Business, and Enterprise plans plus the OpenAI API (model ID gpt-6-astra), Amazon Bedrock, and Microsoft Azure following over the coming days. Capability differs by channel: the standard API and ChatGPT builds refuse advanced offensive tasks such as exploit discovery, while vetted defenders receive broader access through Daybreak Blue. Because the rollout is still staged, this catalog labels the model Preview; it can move to GA once broad availability is confirmed.

Safety classification

GPT-6 Astra is the first model OpenAI has designated “Critical” for cybersecurity under its Preparedness Framework, reflecting its finding that the model can autonomously discover previously unknown security weaknesses and build working exploits against well-defended systems without a human directing every step. OpenAI shipped it with additional safeguards and tiered access, and says it was pre-trained on more than 100,000 GPUs in its largest training run to date.

Migration notes

Set the API model to gpt-6-astra. Workloads that used none or minimal reasoning should start at low. The model does not accept custom temperature, top_p, or top_logprobs; Chat Completions also does not accept logprobs. Use the Responses API for tool calling. Fast mode is unavailable with EU data residency and carries no latency SLA.

Official evaluations

OpenAI's GPT-6 Astra launch page publishes 40 results across nine groups: computer use, professional work, coding, academic tasks, science and health, cybersecurity, alignment, long context, and abstract reasoning. DataLearner records every row and preserves conditions such as OSWorld 2.0 offline partial, ScreenSpot-Pro without tools, HLE with tools, length-adjusted HealthBench Professional, and the two MRCR context ranges. For ARC-AGI-3 this catalog records the ARC Prize–verified 62.7% (Semi-Private set, Standard harness, max effort, $26,098) rather than the 99.9% quoted on OpenAI's launch page. That figure comes from OpenAI's own Provider Adapter harness (high effort, $18,817), which preserves opaque reasoning state between requests and is therefore not comparable across providers. Both are state-of-the-art results; only the comparable one is used in rankings. OpenAI says each value is the best score across evaluated effort settings and that runs used either its research environment or the API, so the rows are labeled “mode unspecified” rather than uniformly described as max effort. Computer-use safety, its AutoReview variant, circumvention, ExploitGym honeypot, and hallucination are lower-is-better metrics.

Official sources

GPT-6 Astra

FAQ

Are GPT-6 and Astra separate models?

No. OpenAI's official product name is GPT-6 Astra and its API ID is gpt-6-astra. DataLearner's former GPT-6 and Astra rumor records now resolve to this single official model identity.

When was GPT-6 Astra released?

The OpenAI API changelog records September 3, 2026 as the release date. Access went first to approved customers in Daybreak, OpenAI's application-only cybersecurity program, followed over the coming days by ChatGPT Plus, Pro, Business, and Enterprise plans and by the OpenAI API, Amazon Bedrock, and Microsoft Azure.

What are the context window, output limit, and knowledge cutoff?

GPT-6 Astra has a 1,050,000-token context window, a 128,000-token maximum output, and an April 30, 2026 knowledge cutoff.

How much does the GPT-6 Astra API cost?

Standard short-context rates per million tokens are $10 input, $1 cached input, $12.50 cache writes, and $50 output. Above 272K input tokens, the full request is billed at $20, $2, $25, and $75 respectively. Batch and Flex are half of Standard, and Fast mode is twice the applicable rate.

Which reasoning effort levels does GPT-6 Astra support?

It supports low, medium, high, xhigh, and max. It does not support none. OpenAI recommends starting with low when migrating requests that previously used none or minimal.

Which inputs and tools does GPT-6 Astra support?

It accepts text and images and returns text. Supported capabilities include streaming, function calling, Structured Outputs, web search, file search, image generation, Code Interpreter, hosted shell, Apply Patch, Skills, computer use, MCP, and tool search. Tool calling requires the Responses API.

What is GPT-6 Astra's parameter count and official benchmark score?

OpenAI has not disclosed total or active parameter counts. Its launch page publishes 40 official evaluation results across nine groups, all now recorded here. Each score is the best result across evaluated effort settings rather than a uniform max-effort run, and five safety or hallucination metrics are lower-is-better. ARC-AGI-3 is the one deliberate exception: this catalog records the ARC Prize–verified 62.7% (Standard harness) instead of the 99.9% on the launch page, which comes from OpenAI's own Provider Adapter harness and is not comparable across providers.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code