AIToday
Large Language ModelsAI Business & IndustrySiliconANGLE AIPublished: Oct 1, 2026, 10:01 JST

Google's Gemini 4 Argon goes to cyber defenders first

Google's Gemini 4 Argon goes to cyber defenders first

3 Key Points

  1. What happened

    Google began rolling out Gemini 4 Argon, which beats Anthropic's Claude Opus 5.5 and OpenAI's GPT-6 Astra on most of Google's benchmarks, to vetted cybersecurity defenders via its Fairwind Program, which opened Sept. 3 and has over 650 organizations.

  2. Why it matters

    Outside Google and Fairwind members, no one can use Argon yet, and Google is holding back full release while it hardens safeguards against cyberattack or weapons misuse, so access, not capability, is the near-term constraint.

  3. What to watch

    Whether Argon's benchmark lead holds once launch pricing ends and it moves to Anthropic's $4 and $20 rates; Google has not set a date for paying developers and Google AI Ultra subscribers.

WHO IT HITSCybersecurity teams at Fairwind member organizations, including CrowdStrike and Palo Alto Networks, can already put Argon to work on vulnerability discovery and remediation, while other security vendors and enterprise defenders remain locked out until Google widens access.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Google is releasing Argon through a deliberately narrow door. The model is going first to the Fairwind Program, which opened Sept. 3 with the smaller Gemini 3.8 Flash Cyber model and has since signed up more than 650 organizations, including CrowdStrike and Palo Alto Networks. Koray Kavukcuoglu, Google's chief AI architect, framed that as a phased approach, and Google says it is still hardening safeguards against misuse for cyberattacks or weapons development. Google is also taking part in the U.S. government's voluntary process for pre-release model access.

The capability story sits in the benchmarks Google published. Argon scored 77.9% on DeepSWE v1.1 for long-horizon software engineering tasks, ahead of Claude Opus 5.5 at 74.2%; AutomationBench showed a wider gap of 51.3% to 42.5%; and VentureBeat counted 12 of 18 benchmarks where Argon led outright. Opus 5.5 still holds Terminal-Bench 4.0, and GPT-6 Astra kept FrontierSWE v2. Google set Argon's output limit at 1 million tokens, up from 64,000 for earlier Gemini models, to match that long-horizon work.

That internal use is already concrete. Google engineers are using Argon agents to move C and C++ code to Rust, including the more than 800,000-line Zircon kernel, and on libgav1 the agents rewrote 32,000 lines as safe Rust, making the decoder run 2.7 times faster than the earlier port with no change to video output. A memory-profiling sweep freed more than 300 tebibytes across Google's data centers. In security, Wiz used Argon in its free Scan for Good program and found a critical flaw in health care software that earlier frontier models had missed.

The open question is timing. Google has not given a date for paying developers and Google AI Ultra subscribers, who are next in line, and Argon's $2 and $10 launch pricing moves to Anthropic's $4 and $20 when it ends. Argon's benchmark lead therefore matters most to the defenders who already have access, while its broader impact hinges on how quickly Google widens the door and on whether the model's safeguards hold up as it does.

FAQ
Who can use Gemini 4 Argon right now?
Only Google's internal teams and members of its Fairwind Program, which opened Sept. 3 with the smaller Gemini 3.8 Flash Cyber model and has signed up more than 650 organizations, including CrowdStrike and Palo Alto Networks.
How much will Gemini 4 Argon cost?
At launch it costs $2 per million input tokens and $10 per million output tokens, with cached input priced 95% lower. It then moves to Anthropic's Opus 5.5 rates of $4 and $20 when launch pricing ends.
How does Argon compare to rival models?
On Google's benchmarks Argon scored 77.9% on DeepSWE v1.1 versus Claude Opus 5.5 at 74.2%, and 51.3% to 42.5% on AutomationBench. Opus 5.5 still leads Terminal-Bench 4.0, and GPT-6 Astra kept FrontierSWE v2.
SiliconANGLE AIRead Original Article

Also reported by AI Watch (Impress), Ars Technica AI, TechCrunch AI, The Verge AI, Top Companies AI

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleJensen Huang wins 2026 Van Fleet Award in New York