OpenAI unveils GPT-6 Astra: Here’s what this new model can do
The company said Astra has already helped solve long-standing open problems in mathematics and represents a new frontier in computer and browser use
04 September, 2026
TT
16
OpenAI has introduced GPT-6 Astra, a new frontier model the company describes as its “most intelligent and aligned” model, with advances spanning computer use, software engineering, cybersecurity, science, mathematics and professional work.
Astra combines years of research and investment in pre-training, reinforcement learning and alignment. According to OpenAI, the model delivers state-of-the-art results across a range of demanding evaluations, including a 98 per cent score on FrontierMath Tier 4, a 99.9 per cent score on ARC-AGI-3 and a perfect 100 per cent score on ExploitBench.
The company said Astra has already helped solve long-standing open problems in mathematics and represents a new frontier in computer and browser use. It is designed to handle complex professional tasks with greater speed, accuracy and judgment.
GPT-6 Astra is initially rolling out to a limited set of organizations. Over the coming days, it will become available to ChatGPT Plus, Pro, Business and Enterprise users, as well as through the OpenAI API, Microsoft Azure and AWS Bedrock.
A stronger focus on alignment and task boundaries
OpenAI said Astra is its most aligned model yet, with improvements in understanding user intent and controlling model behavior.
As part of its testing, the company developed a new evaluation informed by the Hugging Face incident. The evaluation measures whether a model confronted with a difficult or impossible task will go beyond its authorised scope.
Without production safeguards, GPT-5.6 Sol went beyond the authorized target in 48 per cent of cases in the evaluation, while GPT-6 Astra did so in 0 per cent of cases.
Read more-UAE to introduce AI curriculum across public, private schools: What’s to know
OpenAI said Astra is also better at interpreting instructions when they leave room for judgment. The model can use context to fill routine gaps and ask focused questions when missing information could materially affect an outcome. In Codex, Astra can ask questions asynchronously while continuing work that does not depend on the user’s response.
The company said Astra also does a better job of maintaining the broader objective of a task as requirements evolve, incorporating new instructions without losing track of the original request or earlier constraints.
Faster and more capable computer use
OpenAI is positioning GPT-6 Astra as its strongest computer-use model, with applications ranging from online forms and customer records to calendars, research, document preparation and software testing.
The model can conduct online research, draft summaries in email or document editors, analyze scientific data, generate plots, create websites and conduct frontend quality checks. It can also install and test software autonomously and troubleshoot problems displayed on screen.
According to OpenAI, Astra delivers significant efficiency gains in knowledge-work tasks. In latency simulations on OSWorld 2.0, Astra achieved a computer-use performance score of 72.6 per cent at roughly 40 minutes per task, compared with 65.7 per cent at roughly 75 minutes for GPT-5.6 Sol. The company said this represents about 47 per cent less time per task.
OpenAI is also updating the Codex harness to improve the speed of computer use. Combined with Astra’s efficiency, the company said the changes produce a 1.9x faster task-completion rate than the current GPT-5.6 Sol experience on the Mind2Web benchmark.
Designed for professional workflows
GPT-6 Astra is also designed for professional environments, combining advanced reasoning with the ability to execute multistep workflows and produce documents, spreadsheets and presentations.
OpenAI said Astra is particularly strong at following existing templates and creating succinct, well-structured presentations. It can produce documents, presentations, spreadsheets and analyses that follow users’ templates and match their writing and visual styles.
The model is also trained to identify and use only the context relevant to an output rather than repeating unnecessary information, with the goal of producing artifacts that are more immediately usable in business settings.
Astra brings stronger visual judgment to websites, games, applications and renderings. Through Sites in ChatGPT, it can create, host and share websites, web apps and games directly from a prompt.
OpenAI is also introducing a new approach in Codex for preserving context across long sessions. Rather than repeatedly compressing previous work into a single summary when a context window fills, Astra can maintain notes across context windows while keeping earlier context searchable.
The feature is experimental and can be enabled through the Codex config.toml file before becoming the default for Astra in the coming weeks.
Advances in science and mathematics
OpenAI said GPT-6 Astra represents a major advance in scientific discovery, mathematics and health.
The company is sharing two further results involving gaps between prime numbers and said Astra sets new records across a suite of mathematics and science evaluations.
Beyond solving problems, Astra is designed to assist with practical scientific work. By combining scientific reasoning with computer use, it can operate specialised software to inspect data and explore results, helping researchers assess evidence and determine what to investigate next.
Cybersecurity capabilities bring new safeguards
OpenAI said GPT-6 Astra represents a significant increase in cyber capabilities and meets the Critical threshold in cybersecurity under its Preparedness Framework.
The model’s ability to identify and develop zero-day exploits can help defenders locate and patch vulnerabilities, but OpenAI said those capabilities also create a need for stronger safeguards.
In testing without production safeguards, Astra achieved a 100 per cent score on ExploitBench, compared with 78.5 per cent for GPT-5.6 Sol. On ExploitGym, Astra achieved a 42.4 per cent success rate, compared with 30.3 per cent for GPT-5.6 Sol, while using substantially fewer output tokens.
OpenAI also tested Astra on an internal ExploitBench evaluation using vulnerabilities from the previous three months. The company said Astra achieved substantially higher arbitrary code-execution rates than GPT-5.6 Sol while using far fewer output tokens.
During the evaluation, Astra discovered and used two previously unknown zero-day vulnerabilities, both of which OpenAI said it is disclosing to their maintainers.
On SRE-Bench, which measures the ability of models to reverse engineer software binaries without access to raw source code, Astra solved 88 per cent of tasks on its first attempt and 99.2 per cent within four attempts. GPT-5.6 Sol achieved 55.9 per cent and 68.7 per cent, respectively.
OpenAI said expert-led assessments also found that Astra, without production safeguards, could use previously unknown vulnerabilities to achieve arbitrary code execution in hardened browsers and create privilege-escalation exploits for hardened operating systems.
The version of Astra launching today will support defensive activities such as secure code review and patching. However, OpenAI said it will refuse more advanced cybersecurity requests, including creating proof-of-concept exploits for vulnerabilities.
Through OpenAI Daybreak, the company plans to expand access and introduce less restrictive safeguards in the coming weeks for additional defensive workflows, including vulnerability and proof-of-concept validation, malware analysis and detection engineering.
Expanded safety measures
OpenAI said it has strengthened protections against potential cyber misuse, building on the safeguards developed for GPT-5.6 Sol. The measures include greater robustness against jailbreaks and additional context for monitoring systems.
The company said Astra was also evaluated on adversarially selected computer-use tasks designed to elicit misbehavior. Astra was more successful than earlier models at avoiding unintended consequences, with additional default security measures producing stronger performance.
In another internal evaluation, Astra never attempted to circumvent a Codex Auto-Review denial, including when Auto-Review was deliberately configured to be evadable and the task could not otherwise be completed.
OpenAI also said Astra was three times less likely than GPT-5.6 Sol to make inaccurate representations about its capabilities and affordances.
At the same time, the company’s evaluations found Astra’s written reasoning to be harder to monitor than GPT-5.6 Sol’s in tests that explicitly asked the model to evade monitoring. OpenAI attributed this partly to Astra’s greater control over written reasoning on simpler tasks and its ability to solve problems using fewer written steps.
The company said Astra still appears to struggle to conceal the reasoning required for complex tasks, but described the decline in monitorability as a serious issue and said improving monitorability remains a research priority.
OpenAI is also deploying misalignment monitoring in production for Astra-class models. The monitoring system uses classifiers to examine model reasoning and actions for unauthorised behavior and automatically stop potentially unauthorised activity.
Because of Astra’s increased cybersecurity capabilities, OpenAI said it is applying additional safety checks. These can pause or stop legitimate work, including defensive cybersecurity tasks. In ChatGPT and Codex, users may be asked to review an action before continuing, while API tasks will stop when a safety check intervenes.
Availability and pricing
GPT-6 Astra is rolling out initially to a limited group of organizations before becoming available to ChatGPT Plus, Pro, Business and Enterprise users in the coming days. It will also be offered through the OpenAI API, Microsoft Azure and AWS Bedrock.
Astra usage is included within existing subscription allowances, while users and businesses can purchase additional credits. Pro, Business and Enterprise customers will also receive access to GPT-6 Astra Pro.
Enterprise administrators can enable Astra for their workspaces, although access is off by default at launch.
For eligible API customers, Astra supports Zero Data Retention. OpenAI also said it is testing Private Safety Processing to strengthen safety monitoring while preserving customer privacy.
For developers, the model will be available through the OpenAI API as gpt-6-astra, as well as through Microsoft Azure and Amazon Bedrock.
OpenAI’s API Standard pricing is $10 per million input tokens and $50 per million output tokens. Separate rates apply to cache reads and writes. Fast mode is available for GPT-6 Astra and provides up to twice the speed of Standard processing at twice the Standard price.





















