AIToday
Large Language ModelsAI Coding AssistantsAI Business & IndustryArs Technica AIPublished: Oct 1, 2026, 06:00 JST

Google unveils Gemini 4 Argon, still locked to internal testers

Google unveils Gemini 4 Argon, still locked to internal testers

3 Key Points

  1. What happened

    Google announced Gemini 4 Argon, claiming industry-leading performance in coding, knowledge work, and cybersecurity, with a 77.9 percent score on the DeepSWE v1.1 benchmark, above GPT-6 Astra, Fable 5.1, and Opus 5.5.

  2. Why it matters

    Google is signaling it can again compete at the frontier after spending the summer on smaller Flash models, though outside users cannot yet verify the claims.

  3. What to watch

    The benchmark claims remain unverified by outside users, and Google has not announced API pricing yet; watch for that pricing and broader availability to follow limited testing.

WHO IT HITSEnterprise AI teams evaluating coding and cybersecurity models will have to wait for access and pricing before they can compare Gemini 4 Argon against GPT-6 Astra, Fable 5.1, or Opus 5.5.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Google had promised Gemini 3.5 Pro in June, but it spent the summer releasing smaller Flash models instead. The arrival of Gemini 4 Argon marks a return to the frontier for the company, which says its engineers are already using the model internally for tasks such as migrating codebases to Rust and saving memory across data centers. Google has also provided benchmark results, including a 77.9 percent score on the DeepSWE v1.1 software engineering test, which it says beats competing models from other labs.

For now, the model remains in limited testing, and Google has not announced API pricing. The company has confirmed one concrete improvement for future users: a 1 million token output limit, up from 64,000 tokens in prior Gemini models, which Google says will let users complete more daunting tasks in a single step.

The stakes hinge on whether outside developers and businesses will be able to test the model and whether Google's performance claims hold up under independent scrutiny. Until pricing and broader access are announced, the competitive picture remains based on Google's own numbers.

FAQ
When can I use Gemini 4 Argon?
Not yet. Google says the model is still in limited testing, and engineers inside the company are already using it extensively.
What is the output limit for Gemini 4 Argon?
Google has confirmed it will support a 1 million token output limit, up from 64,000 tokens in previous Gemini models.
How does Gemini 4 Argon compare to other models on benchmarks?
Google says it hits 77.9 percent on the DeepSWE v1.1 benchmark, higher than GPT-6 Astra, Fable 5.1, and Opus 5.5.
Ars Technica AIRead Original Article

Also reported by AI Watch (Impress), SiliconANGLE AI, TechCrunch AI, The Verge AI, Top Companies AI

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next article"The White House Accord on Super Intelligence" signed by six AI firms