
What happened
Google added agent-based video analysis to Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite. Instead of sampling every frame per second, the model picks relevant sections on its own.
Why it matters
On Google's 1H-VideoQA and LVBench benchmarks, token usage drops 88% while accuracy improves slightly. This lowers cost for developers and supports tasks like finding anomalies or counting repeated motions in hours of footage.
What to watch
The feature is live through the Gemini API with no extra fee, and it will roll out to all Gemini app users on Flash and Flash Lite soon. Later, it will power "Ask YouTube" on the playback page.
Ask the AI about this article →
Summaries like this, in your inbox every morning.
This update refines Google's earlier move into "agentic vision," which shipped for Gemini 3 Flash in January. That feature let the model write and run Python to edit images, checking results in a loop. Now the same principle applies to video. The model decides which parts of a video to inspect, through frames, audio, or transcript, and retrieves only what it needs, cutting token use by up to 88% on benchmarks like 1H-VideoQA and LVBench.
The practical payoff is accuracy on small details, such as state changes or cuts shorter than one second, which static per-second sampling could miss. Google also reports that on LongVideoBench, Gemini 3.7 Flash leads in overall quality and cost efficiency, suggesting the agent-based approach is not just cheaper but often more precise.
Google says the feature will reach all Gemini app users on Flash and Flash Lite soon, with "Ask YouTube" integration following over the coming months. Because the API charges standard token rates with no extra fee, the cost savings could make sophisticated video analysis more accessible to developers.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Nvidia Corp. CEO Jensen Huang said artificial general intelligence has arrived, following OpenAI's launch of G…

Saudi Arabia's state-backed AI company HUMAIN, led by CEO Tareq Amin, is positioning itself as a neutral hub f…

Alibaba's research division released Qwen-Drive 1.0, an AI model that handles spatial perception, traffic Q&A…

A developer tested whether ChatGPT would judge the same remote-work scenario differently when only the subject…

Google DeepMind ran 100 autonomous LLM agents using Gemini 3.1 Pro on 71 math problems

Fukushima Prefecture ran a proof-of-concept in fiscal 2025 with 100 paid accounts for two generative AI servic…
