GPT-6 Astra: What OpenAI Released, Pricing and Benchmarks
OpenAI announced GPT-6 Astra on September 3, 2026, calling it its most intelligent and most aligned model. It carries a 1,050,000 token context window with up to 128,000 tokens of output, costs $10 per million input tokens and $50 per million output tokens, and adds five reasoning-effort settings aimed at long-horizon agentic work. Access opened first to enterprises in OpenAI's Trusted Access Program, then to the API and to ChatGPT Plus, Pro, Business and Enterprise. Sources: OpenAI, CNBC, Al Jazeera, OpenRouter.
The context
OpenAI announced GPT-6 Astra on September 3, 2026, and the framing was as notable as the model. The company called it the “world’s most intelligent and aligned” AI system, then spent an unusual share of the announcement on what could go wrong with it. That is a change of register for a flagship launch, and it is worth understanding before the specifications.
The specifications are, admittedly, large. Astra reads up to 1,050,000 tokens in a single request and writes back up to 128,000. It costs $10 per million input tokens and $50 per million output tokens, with cached prompt prefixes at $1.00 per million, which is the number that decides whether agent workloads are affordable. It ships with five reasoning-effort settings, so the same model can be run cheap and fast or slow and thorough depending on the task.
On benchmarks OpenAI reports 99.9% on ARC-AGI-3, 97.6% on FrontierMath Tier 4, 72.6% on OSWorld 2.0 and 100% on ExploitBench, ahead of both GPT-5.6 Sol and Anthropic’s Claude Fable 5. The result that actually matters for daily use is quieter: 96.3% on MRCR v2, the long-context retrieval evaluation, against 73.8% for GPT-5.6 Sol. A million-token window is only worth paying for if the model can find things inside it, and that gap is the difference between a marketing number and a usable one.
The rollout was staged rather than open. Companies accepted into OpenAI’s application-based Trusted Access Program, built around cybersecurity work, received Astra first. API access and availability in ChatGPT Plus, Pro, Business and Enterprise followed in the days after, along with distribution through Amazon Web Services. The sequencing is the safety argument in practice: a model that scores 100% on a vulnerability-finding benchmark is very good at exactly the thing defenders and attackers both want, and OpenAI chose to hand it to vetted defenders first.
For anyone deciding whether to move, the useful question is not which model is best but which tasks justify the price. Short prompts with simple outputs still run fine, and far more cheaply, on older models. The cases where Astra earns its cost are the long ones: agents running many steps without supervision, large document sets, code that needs reading before it can be written. Routing by task beats migrating wholesale. Sources: OpenAI, CNBC, Al Jazeera, OpenRouter, llm-stats.
People also ask
True or false?
1,050,000 tokens of input, slightly over the round number everyone quotes, with output capped at 128,000. (OpenRouter)
No. Vetted enterprises in the Trusted Access Program went first, and ChatGPT and API access arrived over the days that followed. (CNBC)
On OpenAI's own benchmarks, yes. Those are the vendor's tests picked by the vendor, so treat the lead as a claim awaiting independent replication. (Al Jazeera)
$10 per million in and $50 per million out puts it at the premium end. The only real discount is $1.00 per million on cached prompt prefixes. (OpenRouter)
100% on ExploitBench, which is exactly why the launch post spent so long on misuse. (OpenAI)
The long-context gap settles it: 96.3% on MRCR v2 against 73.8% for GPT-5.6 Sol is a different model, not a rebrand. (OpenAI)
A ChatGPT Plus subscription reaches it, as does the ordinary API. The Trusted Access Program only governed who got it first. (CNBC)