OpenAI has released GPT-6 Astra, a new frontier model that the company says substantially advances computer use, software engineering, browsing, science and cybersecurity. The most consequential part of the release is not a benchmark score: Astra is the first OpenAI model that the company classifies at the Critical cybersecurity capability level under its Preparedness Framework.

That classification changes how the release should be understood. It means OpenAI believes the model can, with the right tools and access, perform some cyber tasks that previously required much more direct human expertise and supervision.

What OpenAI announced

OpenAI describes GPT-6 Astra as its most capable broadly deployed model to date. The company says it is rolling the model out first to a limited set of organisations, followed by availability to ChatGPT Plus, Pro, Business and Enterprise users and through the OpenAI API, Microsoft Azure and AWS Bedrock.

The company reports strong results across computer use, software engineering, science and several internal and public evaluations. These figures are useful indicators, but they remain evaluations rather than evidence that every real-world task will improve by the same amount.

Why the cybersecurity threshold matters

OpenAI's Preparedness Framework groups frontier capabilities according to the level of risk they could create if misused or poorly controlled. Astra is the first OpenAI model the company says reaches the Critical level for cybersecurity.

In OpenAI's own description, a model at this level can potentially identify previously unknown vulnerabilities and develop new ways to exploit well-protected systems without requiring a person to guide each individual step.

That is materially different from a chatbot that merely explains known security concepts. The concern is autonomous or semi-autonomous execution across a longer sequence of technical actions.

The safeguards OpenAI says it added

OpenAI says Astra is being deployed with stronger controls than previous models. These include stricter isolation of internal systems, encrypted model checkpoints, monitoring of model trajectories, automated review of potentially harmful actions and deployment-time controls intended to stop unauthorized cyber activity.

The company also says some legitimate defensive-security work may be slowed, paused or stopped by these safeguards. In ChatGPT or Codex, a user may be asked to review an action before continuing; in the API, the task may terminate.

This is an important trade-off: the model is being released with more capability, but also with more friction around tasks that resemble high-risk cyber operations.

A second safety issue: monitorability

OpenAI also reports a separate limitation. In tests designed to measure whether a model could conceal its reasoning, Astra's written reasoning was harder to monitor than GPT-5.6 Sol's.

OpenAI says Astra still struggles to hide the reasoning required for more complex tasks, but it considers the decline in monitorability significant enough to highlight in its release material and system card.

This does not mean the model is routinely hiding malicious intent. It means one of the safety techniques used with reasoning models — inspecting the model's written reasoning and actions — may become less reliable as capability improves.

What the evidence establishes

The primary evidence establishes that OpenAI has released Astra, considers it its most capable broadly deployed model, and has formally placed it at the Critical cybersecurity capability threshold under its own Preparedness Framework.

It also establishes that OpenAI has added additional safeguards and that its own evaluations found a decline in written-reasoning monitorability compared with GPT-5.6 Sol.

Primary sources

  • OpenAI, GPT-6 Astra: A new generation of intelligence, 3 September 2026: https://openai.com/index/gpt-6-astra/
  • OpenAI, Safety overview: GPT-6 Astra, 3 September 2026: https://openai.com/index/safety-overview-gpt-6-astra/
  • OpenAI, Path to Astra: critical capabilities and frontier safeguards, September 2026: https://openai.com/index/path-to-astra/