
Google has integrated computer-control capability directly into its Gemini 3.5 Flash AI model, enabling it to see and operate screens across computers, browsers, and mobile devices.
On benchmark tests, the model ranks highly for this task, making it practical for developers to build automated agents for software testing and office work.
The feature includes built-in safeguards and is available now through Google's Gemini API and Enterprise Agent Platform.
What happened
Google integrated "Computer Use" into Gemini 3.5 Flash, allowing the model to see, understand, and interact with computers, browsers, and mobile devices on its own. Previously this capability was only available as a separate Gemini 2.5 model. The feature is now available through the Gemini API and the Gemini Enterprise Agent Platform.
Why it matters
On the OSWorld benchmark, Gemini 3.5 Flash scores 78.4, beating Gemini 3 Flash (65.1) and GPT-5.4 mini (72.1), putting it among the top-performing models for computer interaction tasks. This opens the door for developers to build agents that automate software testing, office tasks, and browser workflows across multiple device types.
What to watch
Google has built in two optional enterprise safeguards—one requiring user confirmation for sensitive actions, and another automatically stopping tasks when indirect prompt injections are detected. The company also recommends sandboxing, human oversight, and strict access controls to guard against abuse.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
AI is now helping solve longstanding math problems, including the so-called Jacobian conjecture, and top mathe…

A benchmark tested 15 language models on two briefs requiring cited sources

AMD introduced ROCm 10, the next version of its software stack for AMD Instinct GPUs, designed to streamline t…

Lam Research has started construction on an R&D center in Tualatin, Oregon, as part of a $3 billion, five-year…

Visa says its AI 'harness' makes Anthropic cheaper to use for cyber defense

Visa announced enhancements to its cybersecurity portfolio, including the next evolution of its open-source Vi…
