AIToday
Large Language ModelsAI Business & IndustryLatent SpacePublished: Sep 4, 2026, 16:00 JST2 min read

OpenAI launches GPT-6 Astra, its biggest model yet

OpenAI launches GPT-6 Astra, its biggest model yet

Key takeaway

  • OpenAI launched GPT-6 Astra, its flagship model. It claims state-of-the-art computer use and software engineering.

  • The launch broke popularity records, with 36M views in hours.

  • Benchmarks show progress but not universal dominance over rivals like Fable.

3 Key Points

  1. What happened

    OpenAI launched GPT-6 Astra as its new flagship model, calling it "our most intelligent and aligned model yet." The rollout was bumpy, with delays and unclear access timing, leading the company to offer "banked resets" for paid users lacking access.

  2. Why it matters

    This launch marked the first time OpenAI outpaced Anthropic in launch popularity, with 36M views and 164K likes within 9 hours. Benchmark results were mixed: Astra tied Claude Opus 5 and Fable 5 on the Coding Agent Index (67) but scored 5 points lower than Fable 5.1 on intelligence (61), sparking debate over whether gains were uneven or a true step-change.

  3. What to watch

    Pricing is $10/$50 per 1M input/output tokens standard, and $20/$100 fast. Astra is rolling out first to limited organizations, then to Plus/Pro/Business/Enterprise, API, and AWS over coming days. Notable third-party results include 99.9% on ARC-AGI-3 with a provider harness and 46.7% on MirrorCode, between Opus 4.7 and Fable 5.

Ask the AI about this article →

Context & Analysis

The launch of GPT-6 Astra comes amid intense rivalry with Anthropic, whose models Fable and Opus have recently set high bars. OpenAI's positioning emphasizes computer use, software engineering, math/science, and cybersecurity, but external evaluations tell a more nuanced story. For instance, Artificial Analysis reports Astra ties on coding but lags on intelligence, while ARC Prize highlights a breakthrough with a crucial harness caveat, noting performance jumps from 63% to 99% depending on the setup.

The mixed results have fueled skepticism about benchmark saturation and evaluation-awareness. Some researchers argue visible alignment gains may be "papering over" specific failure modes rather than solving underlying goal misalignment. The system card's discussion of decreased chain-of-thought monitorability adds to concerns, as does UK AISI's finding that Astra can operate without reasoning for far longer than predecessors, potentially evading monitoring.

Despite the debate, the launch's popularity marks a shift in public perception. OpenAI's success with 36M views in under a day suggests that, for the first time, it has matched or exceeded Anthropic's ability to generate excitement. The company's "banked resets" for delayed paid users indicate an attempt to manage expectations, while the staged rollout to organizations before broader access reflects a cautious approach to deployment. As benchmarks saturate quickly—ARC-AGI-4 is coming Q1 2027—the focus may shift from raw capability to cost efficiency and real-world utility, where Astra shows promise but not clear supremacy.

FAQ

How much does GPT-6 Astra cost?
Standard pricing is $10 per 1M input tokens and $50 per 1M output tokens. Fast mode is $20 per 1M input and $100 per 1M output, offering up to 2.5x speed.
What are the most controversial findings in the system card?
The system card describes decreased chain-of-thought monitorability, meaning OpenAI has less visibility into the model's reasoning. UK AISI measured Astra's no-CoT time horizon at 30.9 minutes versus 3.6 minutes for GPT-5.6 Sol.
How does Astra perform on coding benchmarks?
On the Coding Agent Index, Astra scored 67, about equal to Claude Opus 5 and Fable 5, while Fable 5.1 leads with 70. Astra is 70% more token efficient than GPT-5.6 Sol and uses one fifth the tokens of Claude Opus 5.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Infosys and CrowdStrike partner on AI-discovered vulnerabilitiesSiliconANGLE AI · 1h ago
  • Cisco sets zero-engineers-coding targetDIGITIMES Asia · 1h ago
  • OpenAI's Astra model thinks beyond human oversightSemafor Tech · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleGrok 4.6 Now in Public Preview on Snowflake Cortex AI