Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

https://storage.googleapis.com/gweb-uniblog-publish-prod/images/gemini-3-5_3-6_3-5-Cyber__key-art__statement_.width-1300.jpg

Jul 21, 2026

Our newest Gemini models deliver the efficiency, latency, and reliability to build AI agents at scale.



Listen to article

[[duration]] minutes

This content is generated by Google AI. Generative AI is experimental

Developers and customers building production AI agents need higher token efficiency, lower latency, and more reliable performance. Our Flash series of models is built to meet the sweet spot of efficiency and quality to enable scaling agentic workflows. Building on Gemini 3.5 Flash, we’re introducing new Gemini models:

  • 3.6 Flash: Our workhorse model that delivers better coding, knowledge work, and multimodal performance. According to the Artificial Analysis Index, it reduces output token usage by 17% compared to 3.5 Flash, and in some benchmarks like DeepSWE by Datacurve, we observe up to 65%, all at a lower cost per output token.
  • 3.5 Flash-Lite:Our fastest, most cost-effective 3.5-class model, delivering 350 output tokens per...

Copyright of this story solely belongs to blog.google. To see the full text click HERE

Read more