Google released three new Gemini models —3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber— aimed at efficiency, low latency, and reliability for those building AI agents at scale. Notably, it did not include the anticipated Gemini 3.5 Pro, apparently delayed over performance issues.
The 3.6 Flash is the “workhorse” model, with better code and up to 17% lower token use than its predecessor, which makes it cheaper. The Flash-Lite is the most economical in its class, for high-volume automation, and the Flash Cyber specializes in finding and fixing vulnerabilities (limited access, in pilot).
Why it matters to you
Cheaper, faster models lower the cost per task. That lets you run high-volume automations —sorting messages, answering inquiries, following up— without the spend spiking as you grow.