Google has released the Gemini 2.5 Computer Use model, a specialized model that enables agents to interact with user interfaces, such as web browsers and mobile applications. The model is available via the Gemini API and can be used to automate tasks that require direct interaction with graphical user interfaces. It outperforms leading alternatives on multiple web and mobile control benchmarks with lower latency. The model can be used to automate tasks such as filling out forms, manipulating interactive elements, and operating behind logins.