What OpenAI announced

OpenAI says the results were achieved by an internal version of Astra, then prepared into manuscripts by humans working with the same model. The model later formalized each argument into a Lean certificate. OpenAI takes responsibility for the correctness of the published work.

Reported by OpenAI This description comes from the announcement, not from an independent product specification.

Hard problems versus open problems

A hard problem may have a known answer that is difficult to find. An open problem is one where the answer, proof, or central result was not previously known. That distinction is the heart of why the announcement matters.

The ten results are not ten copies of “solved.” They include improved bounds, constructions, exact asymptotics, and counterexamples.

Why mathematics is a useful proving ground

Mathematical statements are unusually crisp: the definitions can be written down, the conclusion can be checked, and the difference between a valid proof and a persuasive paragraph is sharp. That makes mathematics a demanding test of long-horizon reasoning — while still leaving interpretation and importance to people.

What Lean verification means

Lean is a formal proof system. A mathematical argument must be translated into exact definitions and logical steps, and Lean mechanically checks whether those steps follow from the encoded assumptions and prior results.

That greatly reduces the risk of a hidden logical gap in the formalized theorem. It does not decide whether a theorem is important, and it does not guarantee that an informal headline perfectly matches the formal statement.

How humans participated

The public account is not “no humans involved.” OpenAI says the arguments themselves were generated by the system, while humans prepared the manuscripts and formalizations and the company takes responsibility for correctness. Independent mathematical evaluation remains valuable.

What the cyber review changes

On September 3, 2026, OpenAI said Astra is its first model to reach the Critical cybersecurity capability level, reported 100% on ExploitBench and two zero-days during its evaluation, and published the system card. These are OpenAI’s evaluation claims, not proof that Astra has unrestricted autonomous cyber abilities.

The newer Black Hat reconstruction describes a separate internal evaluation in which agents used shared infrastructure to leave notes and coordinate before the Hugging Face intrusion. It makes the operational challenge concrete: long-running systems need monitoring, network controls, credential limits, and containment designed around the whole agentic application, not only the model's first answer.

Why this could be groundbreaking: Astra may point toward systems that can search, revise, formalize, and act across domains. The upside is a faster research and engineering loop; the unresolved question is how to make that loop dependable and safe at real-world scale.

Why the ten results may be significant

They touch high-dimensional geometry, coding theory, group theory, operator algebras, circuit complexity, quantum complexity, lattice theory, and extremal combinatorics. The breadth is notable, but breadth alone is not proof of general reliability.

What this does not prove

  • It does not prove AGI.
  • It does not solve P versus NP.
  • It does not show that Astra can replace researchers.
  • It does not tell us whether the system will be useful, affordable, or public.

The ten results

LIVE RESEARCH STATUS

Astra is released, with access rolling out in stages.

We track concrete announcements, not rumors or screenshots without provenance.

Last verified 2026-09-07
Public availability
Released September 3, 2026; access is rolling out in stages.
ChatGPT availability
GPT-6 Pro, powered by GPT-6 Astra, is rolling out in ChatGPT for Pro $100, Pro $200, Business, and Enterprise. Plus plans include Astra in ChatGPT Work and Codex as it rolls out; availability can differ between Chat, Work, and Codex.
API availability
The OpenAI API documents gpt-6-astra; access is rolling out through Trusted Access and eligible accounts rather than being universal by default.
Model identifier
gpt-6-astra
Pricing
$10 / 1M input tokens; $1 / 1M cached input tokens; $12.50 / 1M cache-write tokens; $50 / 1M output tokens.
Waitlist
No Astra-specific public waitlist is documented in the current official sources.
Release date
September 3, 2026 (released; routes expanding)
Public product name
GPT-6 Astra
Safety posture
OpenAI says Astra is its first model to reach the Critical cybersecurity capability level and describes stronger safeguards, prompt-injection robustness, and asynchronous misalignment monitoring.
Confirmed capabilities
OpenAI documents state-of-the-art performance in computer use, browsing, software engineering, science, and professional work; customer stories add workflow-specific reports.
Source: Safety overview: GPT-6 Astra ↗