
What happened
Google released Gemini 3.6 Flash to replace 3.5 Flash, along with Gemini 3.5 Flash Lite and Gemini 3.5 Flash Cyber. The company did not release the delayed Gemini 3.5 Pro, which was supposed to launch in June.
Why it matters
3.6 Flash improves coding performance (49% on DeepSWE test versus 37% for 3.5 Flash) and uses about 17 percent fewer tokens while lowering API costs to $1.50/1M input tokens and $7.50/1M output tokens (down from $1.50 and $9 for 3.5 Flash). These efficiency gains help developers and businesses reduce AI token costs.
What to watch
Flash Lite now processes at 350 tokens per second and costs $0.30/1M input tokens and $2.50/1M output tokens, positioning it as Google's most efficient modern AI. The Gemini 3.5 Pro release date remains unclear.
Summaries like this, in your inbox every morning.
Google's decision to skip the Gemini 3.5 Pro release signals a strategic shift toward efficiency and cost control. The company emphasized that changes to 3.6 Flash were made in response to user feedback on 3.5, particularly around code generation performance. Google's focus on token efficiency reflects broader industry pressure: as businesses have started to fret over the cost of AI tokens, reducing compute requirements while maintaining capability has become a competitive priority. The improved coding performance (49% versus 37% on DeepSWE) combined with 17 percent fewer token usage demonstrates that Google has addressed earlier shortcomings without adding cost burden. The introduction of Flash Lite at 350 tokens per second positions Google to compete in price-sensitive, agentic workflows where cost-per-inference is the limiting factor. The delayed Pro version suggests Google is prioritizing the optimized models that better serve existing demand rather than rushing a higher-tier release.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, LINE, or Slack.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Xiaomi released and open-sourced the MiMo-V2.6 series — MiMo-V2.6-Pro and a smaller Flash variant — plus a Pro…
Hermes Testing Solutions began trading on the over-the-counter market on September 22, aiming to benefit from…

The Indeed Hiring Lab report says pay in the most AI-exposed US occupations rose roughly 46% since 2021, versu…

At Semafor's The Next 3 Billion event, Nvidia sustainability head Josh Parker attributed recent US anti-AI sen…

JS Denain of Epoch AI said OpenAI and Anthropic blog posts on AI accelerating AI progress are not strong evide…

The Biological Computing Co
