AIToday
Large Language ModelsAI Safety & AlignmentOpenAI BlogPublished: Sep 10, 2026, 10:00 JST2 min read

GPT‑6 Astra cuts unintended outcomes 89% vs GPT‑5.6 Sol

GPT‑6 Astra cuts unintended outcomes 89% vs GPT‑5.6 Sol

3 Key Points

  1. What happened

    OpenAI introduced GPT‑6 Astra, now in ChatGPT Work, Codex, and the API. On its internal computer use safety benchmark, Astra produced unintended outcomes 89% less often than GPT‑5.6 Sol and 74.7% less often than Claude Fable 5.1.

  2. Why it matters

    Most AI requires businesses to prepare data and build integrations first. Astra works through everyday applications even without an API, so teams can deploy within existing workflows from day one rather than after engineering work.

  3. What to watch

    Astra is the first model to reach the Critical cybersecurity threshold under OpenAI's Preparedness Framework, so the test is whether its strengthened protections hold up as access expands. Watch the $10 per million input tokens and $50 per million output tokens pricing.

WHO IT HITSEnterprise IT and security teams evaluating AI for sensitive workflows will need to weigh Astra's reduced unintended-outcome rate against its Critical cybersecurity designation and the new admin controls for restricting access.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

OpenAI positions Astra as a model that fits into existing workflows rather than requiring businesses to rebuild them. The company says Astra can write code and work through the same applications people use daily, even when those applications lack an API. It also highlights early customer uses, from optimizing GPUs to spotting discrepancies in financial statements, and internal tests such as turning three hours of multicamera footage into a video that drew over 550k views in four days.

The launch pairs capability gains with control features. Astra is the first model to reach the Critical cybersecurity capability threshold under OpenAI's Preparedness Framework, and OpenAI says it strengthened protections against misuse and unauthorized actions. At the same time, new enterprise admin controls let organizations limit access to approved websites and apps, manage uploads and downloads, and require approval before consequential actions.

For teams weighing adoption, the outcome may hinge on whether those safeguards and controls work as intended when Astra is given access to consequential business systems. The stated safety results and the Critical designation point in opposite directions, so the practical test is whether organizations can start with a limited configuration and expand access without unexpected incidents.

FAQ
How much does GPT‑6 Astra cost?
Pricing starts at $10 per million input tokens and $50 per million output tokens.
What can businesses control about Astra's access?
New enterprise admin controls let organizations restrict access to approved websites and desktop applications, manage uploads and downloads, and control browsing history.
Which benchmark shows Astra's safety improvement?
On OpenAI's internal computer use safety benchmark, Astra produced unintended outcomes 89% less often than GPT‑5.6 Sol and 74.7% less often than Claude Fable 5.1.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • DeepSeek V4.1-Flash: 763B model beats V4 Pro on AA Index 40Latent Space · 1h ago
  • Dynatrace acquires Arize AI as observability shifts to actionSiliconANGLE AI · 7h ago
  • Shared base cuts 100 fine-tunes from 1.5 TB to 19.3 GBDaily Dose of Data Science · 7h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleUnitedHealth to Invest $1.5B in AI to Boost Optum Insight