GPT-6 Astra
Released
GPT-6 Astra is an AI model released by OpenAI on Thursday, Sep 3 2026, 24 days after GPT-5.6-Cyber. It has a 1.05M token context window. Benchmark results (shown below) cover Auto-review circumvention (Internal), FrontierCode v1.1 (Main), AA Coding Agent Index, Database Migration Tasks (OpenAI Internal), BenchCAD, Terminal-Bench-Science 0.1, and 17 more.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $10.00
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $1.00
- Cache writeA one-off charge for storing text so later requests can reuse it at the cheaper cached rate.
- $12.50
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $50.00
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $20.00
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $2.00
- Cache writeA one-off charge for storing text so later requests can reuse it at the cheaper cached rate.
- $25.00
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $75.00
- Reasoning or thinking is supported.
- OpenAI also publishes a Fast mode for this model at exactly twice the standard rate in every column — $20.00 per million input tokens and $100.00 per million output at short context.
Benchmarks
Coding
Terminal & CLI
Agentic & tool use
Reasoning & science
Knowledge work
Healthcare
Multimodal
Robustness
About
GPT-6 Astra, released September 3, 2026, opened OpenAI's GPT-6 generation and was the company's first new general-purpose flagship since the three-model GPT-5.6 family — Sol, Terra and Luna — arrived on June 26, 2026. OpenAI had trailed the model for a month before launch without ever announcing it as a product: on August 1, 2026 it named Astra "our next major model" in a research post, reporting that an internal version had produced new results on ten long-standing open mathematical problems, each with a machine-checkable proof in the Lean theorem prover. Capability demonstrated in a paper before a model existed to sell was an unusual way to introduce a flagship, and it set the terms for the four weeks that followed.
The launch table led on reasoning: 99.9% on ARC-AGI-3, against the 30.2% Claude Opus 5 had posted on the same interactive-environment board in July, 97.6% on the rebuilt second edition of FrontierMath Tier 4, and 96.0% on GPQA Diamond. Agentic work moved by similar margins — 64.6% on Terminal-Bench Science 0.1 and 41.4% on AutomationBench, both more than double GPT-5.6 Sol's figures in the same run, and 74.1% on DeepSWE v1.1. The security rows are the ones that explain the wait: 100.0% on ExploitBench, the Carnegie Mellon capability ladder that scores how far a model gets turning a known V8 bug into a working exploit, and 99.2% on SRE-Bench over four attempts.
The professional and coding tables ran closer. Astra took Terminal-Bench 4.0 at 57.9% against Claude Fable 5.1's 55.8%, BrowseComp at 91.5%, and 0.84 on the OpenScore String Quartets sheet-music transcription set where GPT-5.6 Sol had managed 0.19 — but on FrontierCode 1.1 its 53.3% on the main split and 64.5% on the extended set both sat fractionally behind the Claude models OpenAI ran beside it, and it trailed on Artificial Analysis' two composite indices: 61.2 on the Intelligence Index against Fable 5.1's 65.7, and 67.0 on the Coding Agent Index. For a generational release the shape of the table was lopsided — enormous margins on reasoning, security and the internal professional sets, and a coding field the previous generation had already crowded.
That capability is what OpenAI spent August evaluating. On August 7 the company said it had paused parts of the work on Astra because preliminary results could not rule out that the model reached the Critical cybersecurity level in its Preparedness Framework — the threshold describing a model able to find and carry out attacks against well-defended systems without human help, and the first time OpenAI had put one of its own models in that bracket. It returned to the question on September 1 in "Path to Astra: critical capabilities and frontier safeguards", saying the further evaluations were complete and the safeguards it had built around the model were sufficient to ship. Astra launched two days later.
Compare GPT-6 Astra with
Suggested comparisons
Frequently asked questions
GPT-6 Astra was released by OpenAI on Thursday, Sep 3 2026.
GPT-6 Astra was built by OpenAI. Creators of ChatGPT and the GPT series of models. Pioneered large-scale language model research.
GPT-6 Astra costs $10.00 per million input tokens and $50.00 per million output tokens through the OpenAI API. Cached input is $1.00 per million tokens. Those are the rates for the “Up to 272K input tokens” tier; 1 other pricing tier is published for this model. Rates are pay-as-you-go API prices verified against OpenAI's published pricing on September 3, 2026.
GPT-6 Astra reports 24 tracked benchmark scores — Auto-review circumvention (Internal): 0%; DeepSWE 1.1: 74.1%; FrontierCode v1.1 (Main) (main split): 53.3%; FrontierCode v1.1 (Extended) (extended split): 64.5%; AA Coding Agent Index: 67; Database Migration Tasks (OpenAI Internal): 63.9%; BenchCAD: 95.9%; Terminal-Bench 4.0: 57.9%; Terminal-Bench-Science 0.1: 64.6%; BrowseComp: 91.5%; ExploitBench: 100%; ARC-AGI-3: 99.9%; FrontierMath (Tier 4 (v2)): 97.6%; GeneBench-Pro: 39%; MedChemBench (Internal): 49.7%; GPQA Diamond: 96%; Agent's Last Exam (pass@1): 59.3%; SRE-Bench: 99.2%; AutomationBench: 41.4%; HealthBench Professional (length-adjusted): 63.4%; AA Intelligence Index: 61.2; Design Tasks (OpenAI Internal): 50%; Data Science Tasks (OpenAI Internal): 40.9%; OpenScore String Quartets: 0.84. Scores are the figures published at release by OpenAI. It holds the best score among all models tracked here on Auto-review circumvention (Internal), FrontierCode v1.1 (Extended) (extended split), AA Coding Agent Index, Database Migration Tasks (OpenAI Internal), BenchCAD, Terminal-Bench-Science 0.1, ExploitBench, ARC-AGI-3, FrontierMath (Tier 4 (v2)), GeneBench-Pro, MedChemBench (Internal), GPQA Diamond, Agent's Last Exam (pass@1), SRE-Bench, HealthBench Professional (length-adjusted), AA Intelligence Index, Design Tasks (OpenAI Internal), Data Science Tasks (OpenAI Internal) and OpenScore String Quartets.
GPT-6 Astra has a context window of 1.05M. That is the maximum amount of input plus output the model can hold in a single request.
No. GPT-6 Astra is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
OpenAI's previous tracked release was GPT-5.6-Cyber on Aug 10 2026, 24 days earlier. It is the most recent OpenAI model tracked on AI Release Tracker.