
Google released three new Gemini models on Tuesday—Gemini 3.6 Flash with 17% lower token usage, cost-effective 3.5 Flash-Lite, and cybersecurity-focused 3.5 Flash Cyber for governments and partners. The update notably omits the long-awaited Gemini 3.5 Pro, which was last updated in February and has faced internal delays as OpenAI and Anthropic accelerate their own model releases.
Summaries like this, in your inbox every morning.
Sign up free →What happened
Google DeepMind released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber on Tuesday. Gemini 3.6 Flash reduces token usage by up to 17% compared to its predecessor 3.5 Flash while improving coding and multimodal capabilities. 3.5 Flash Cyber, a cybersecurity-focused model, will be exclusively available to governments and trusted partners through a limited access pilot program.
Why it matters
The launch emphasizes efficiency and speed for AI agents running at scale, addressing a core need for production systems. However, Google notably did not release the long-anticipated Gemini 3.5 Pro update—last updated in February—signaling internal delays as competitors including OpenAI (which has released GPT-5.5 and begun rolling out GPT-5.6) and Anthropic (which launched Claude Opus 4.8, Claude Sonnet 5, and expanded Fable 5 access) maintain a faster release cadence.
What to watch
Google DeepMind product lead Logan Kilpatrick said Tuesday the company is currently testing Gemini 3.5 Pro with partners and hopes to "land soon." The team has also started its most ambitious pre-training run yet for Gemini 4, according to Kilpatrick.
On Tuesday, Google DeepMind released three new Gemini models: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber. Gemini 3.6 Flash is positioned as Google's "workhorse model" and offers improved capabilities in coding, knowledge work, and multimodal performance. A key efficiency gain is a reduction in token usage by up to 17% compared to its predecessor 3.5 Flash, lowering inference costs. Gemini 3.5 Flash-Lite is the most cost-effective model in its class, while Gemini 3.5 Flash Cyber is a specialized, fine-tuned model designed specifically for finding and fixing cybersecurity vulnerabilities. Google announced that 3.5 Flash Cyber will be exclusively available to governments and trusted partners through a limited access pilot program.
The company emphasized that these releases prioritize efficiency, latency, and reliability for customers building AI agents at scale. Flash models generally prioritize lower cost and faster response times for production applications, in contrast to Pro models, which are optimized for complex reasoning and coding tasks and represent Google's highest-capability offerings.
The update was notable as much for what Google did not release as for what it did. The long-anticipated Gemini 3.5 Pro update did not ship; Pro was last updated in February. Google had teased the Pro release in May alongside the 3.5 Flash announcement, stating the Pro version was "already being used internally, and we look forward to rolling it out next month." Last week, Bloomberg reported that Google faced internal delays launching Gemini 3.5 Pro as the company struggled to meet internal performance goals. In the same period, OpenAI has released GPT-5.5 and begun rolling out GPT-5.6, while Anthropic launched Claude Opus 4.8 and Claude Sonnet 5 and expanded access to its frontier Fable 5 model, highlighting the intense release pace of rival AI labs.
Google DeepMind product lead Logan Kilpatrick acknowledged the delay on Tuesday, saying the company is currently testing Gemini 3.5 Pro with partners and hopes to "land soon." He also noted that the team has started its most ambitious pre-training run yet for Gemini 4, signaling continued development of next-generation capabilities.
Google's Tuesday release reflects a strategic pivot toward efficiency and operational scale at a moment when its competitors are accelerating their own cadence. While the company shipped three models—each addressing a specific use case (cost-effectiveness, faster response, and cybersecurity)—the absence of Gemini 3.5 Pro reveals internal constraints. The article cites Bloomberg's reporting of performance-goal misses, suggesting Google's engineering timeline has slipped. In the interim, OpenAI has moved beyond GPT-5.5 to begin rolling out GPT-5.6, and Anthropic has launched multiple frontier models including Claude Opus 4.8 and expanded access to Fable 5. This competitive backdrop frames Google's focus on cheaper, faster inference as a near-term strategy: serving production customers at scale while flagship-model development continues internally. Logan Kilpatrick's statement that the team has begun "its most ambitious pre-training run yet for Gemini 4" signals that Google's long-term roadmap remains ambitious, even as near-term releases lag rival labs' pace.
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
No comments yet. Be the first to share your thoughts!
Log in to join the discussion




Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.
Get Started FreeFree · takes 30 seconds · unsubscribe anytime
1 minute a day. The AI essentials.
200+ sources · Email / LINE / Slack