AIToday
Large Language ModelsAI Business & IndustryAI Safety & AlignmentSiliconANGLE AIPublished: Sep 4, 2026, 10:00 JST2 min read

OpenAI rolls out GPT-6 Astra, perfect on 3 benchmarks

OpenAI rolls out GPT-6 Astra, perfect on 3 benchmarks

Key takeaway

  • OpenAI has started rolling out GPT-6 Astra. It scored perfectly on three top AI benchmarks.

  • Astra's hacking abilities delayed its release for safety.

  • OpenAI is giving $1 billion in credits to essential service operators.

3 Key Points

  1. What happened

    OpenAI started opening access to GPT-6 Astra, its newest and most capable large language model. The model scored 98% on FrontierMath Tier 4, 99.9% on ARC-AGI-3, and a perfect score on ExploitBench.

  2. Why it matters

    OpenAI says Astra shows "state of the art" performance in coding, browsing, and computer use. Its cybersecurity skills led OpenAI to label it a "critical" risk, delaying release for several weeks while engineers developed guardrails.

  3. What to watch

    At launch, Astra is only available to a limited number of customers via the Daybreak program. OpenAI will provide $1 billion in credits to help government agencies, utilities, and other essential service operators improve cybersecurity.

Ask the AI about this article →

Context & Analysis

GPT-6 Astra's release follows strong pre-release results, including solving several Erdos problems and making advances in computational complexity theory. Its performance on benchmarks like ARC-AGI-3, which measures learning ability, and ExploitBench, which tests vulnerability exploitation, highlights its advanced capabilities. The "critical" risk designation under OpenAI's internal safety framework underscores the potential dangers of such powerful models, leading to a delay to implement necessary safeguards.

OpenAI attributes Astra's complex task handling to a new data management approach. Unlike previous models that compress prompts and discard low-priority information, Astra archives it in a searchable form. This retention of data points could improve output quality, suggesting a shift in how models manage memory and context.

The $1 billion in credits for Daybreak participants, provided in partnership with MS-ISAC for training, signals a strategic focus on using Astra for cybersecurity defense. This move may help essential service operators bolster their defenses against cyberattacks, though the model's own offensive capabilities remain a concern. The gradual rollout to ChatGPT, Codex, and the API will likely expand Astra's impact across various applications.

FAQ

When will GPT-6 Astra be available to everyone?
Astra is initially available only to a limited number of customers via the Daybreak program. OpenAI will expand availability over the coming days by bringing it to ChatGPT, Codex, and its application programming interface.
Why was GPT-6 Astra's release delayed?
OpenAI disclosed that Astra qualifies as a "critical" risk under its internal AI safety evaluation framework because it can hack "many well-protected systems" without human input. Engineers spent several weeks developing guardrails against hacking.
SiliconANGLE AIRead Original Article

Also reported by ITmedia AI+, WIRED AI

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Nvidia buys Hugging Face for $12.93B, pays rivals, retains staffDIGITIMES Asia · 1h ago
  • Nvidia buys Hugging Face for US$12.93 billionDIGITIMES Asia · 1h ago
  • AI worm spreads via Word docs in CopilotITmedia AI+ · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleNvidia buys Hugging Face for $12.93B, pays rivals, retains staff