Google has introduced new Gemini models, including 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber, which aim to deliver improved efficiency, latency, and reliability for building AI agents at scale. The 3.6 Flash model is said to be 17% more token-efficient than its predecessor and performs better in coding and tool usage benchmarks. The models are designed to meet the needs of developers and customers building production AI agents. The company is also working on the next generation of models, including Gemini 4.