Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Google has announced the release of new Gemini 3.6 Flash models, focusing on improved token efficiency, lower latency, and reduced costs for AI agent development. The company is also currently testing Gemini 3.5 Pro and has begun pre-training for Gemini 4.
Why it matters
Advancements in model efficiency are essential for scaling AI agent workflows in production environments, making AI more accessible and cost-effective.
Our newest Gemini models deliver the efficiency, latency, and reliability to build AI agents at scale.
Senior Director, Product Management, on behalf of the Gemini team
Your browser does not support the audio element.
Developers and customers building production AI agents need higher token efficiency, lower latency, and more reliable performance. Our Flash series of models is built to meet the sweet spot of efficiency and quality to enable scaling agentic workflows. Building on Gemini 3.5 Flash, we’re introducing new Gemini models:
Beyond today’s releases, Gemini 3.5 Pro is currently testing with partners and we plan to make it broadly available as soon as it’s ready. In parallel, our team is already focusing on building the next generation of models. We have started our most ambitious pre-training run yet, for Gemini 4, and are excited by the progress.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in