More on the topic...
Generating detailed summary...
Failed to generate summary. Please try again.
Gemini 3.5 Flash now includes a built-in computer use tool, bringing the standalone Gemini 2.5 computer use features directly into the main Flash model. That means the same model you call for chat and function-calling can now automate tasks across browsers, mobile apps and desktop environments. Google says this upgrade boosts performance on long-horizon workflows like continuous software testing and knowledge work in enterprise apps.
To tap into the new feature, developers use either the Gemini API or the Gemini Enterprise Agent Platform. In practice, you can ask 3.5 Flash to scan an application’s UI and list out its menu items or audit your own docs for accessibility problems. Under the hood, Google applied targeted adversarial training to limit prompt-injection attacks when agents operate in live settings.
On top of that, two enterprise safeguards are optional: one forces the agent to pause and ask for user approval before any sensitive or irreversible action; the other automatically stops a task if it spots a suspicious prompt pattern. Google recommends pairing these controls with sandboxing, human review steps and tight access management. If you want to try it yourself, there’s a demo on Browserbase and full reference code in the Gemini documentation.
Questions about this article
No questions yet.