AIToday
Large Language ModelsLatent SpacePublished: May 20, 2026, 13:01 JST

Google launches Gemini 3.5 Flash with 1M-token context and Gemini Omni video generation at I/O 2026; processes 3.2 quadrillion tokens/month across 900M+ monthly users

Google launches Gemini 3.5 Flash with 1M-token context and Gemini Omni video generation at I/O 2026; processes 3.2 quadrillion tokens/month across 900M+ monthly users

3 Key Points

  1. Gemini 3.5 Flash is now generally available with 1M-token context window, 65k max output tokens, four thinking levels (minimal, low, medium, high), and thought preservation across multi-turn conversations; Google claims it runs 4x faster than comparable frontier models and up to 12x faster in Antigravity (a platform for running background agents and long-horizon tasks).

  2. Gemini Omni, a new family combining Gemini reasoning with generative media capabilities, takes text/image/video/audio inputs and produces video edits and generation in Gemini, Flow, Shorts, and later via APIs; Omni Flash is available in Gemini and Flow today for paid users and in Shorts and Create starting this week for free users.

  3. Google reports processing 3.2 quadrillion tokens/month, up 7x year-over-year from 480 trillion/month; Gemini app has 900M+ monthly users and is available in 230+ countries and 70+ languages. Artificial Analysis benchmarks show Gemini 3.5 Flash at Intelligence Index 55 (+9 vs. Gemini 3 Flash), GDPval-AA 1656 Elo, and pricing of $1.50 / $9.00 per 1M input/output tokens; independent Arena reports the model at #9 in Text Arena and #9 in Code Arena: Frontend with a score of 1507 (+70 over Gemini 3 Flash).

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleGoogle redesigns search and launches AI agents across products as it plans to spend $180 billion to $190 billion on AI infrastructure this year