Gemini last models: temperature, top_p, and top_k are deprecated and ignored

Google has released Gemini 3.6 Flash and 3.5 Flash-Lite, while announcing that older parameters like temperature and top_p are now deprecated. The update includes new pricing structures and integration tools for developers.
Why it matters
These changes impact the development workflows of thousands of applications relying on Google's AI infrastructure, necessitating code migrations.
Gemini 3.6 Flash ( gemini-3.6-flash ) and Gemini 3.5 Flash-Lite ( gemini-3.5-flash-lite ) are generally available (GA) and ready for production use.
This guide explains what's new in each model, what API changes affect your code, and how to migrate.
npx skills add google-gemini/gemini-skills --skill gemini-interactions-api --global Apply the skill:
/gemini-interactions-api migrate my app to Gemini 3 .6 Flash Gemini 3.5 Flash-Lite Install the skill:
/gemini-interactions-api migrate my app to Gemini 3 .5 Flash-Lite New models Model Model ID Default thinking level Pricing Description Gemini 3.6 Flash gemini-3.6-flash medium $1.50/1M input tokens and $7.50/1M output tokens Balances speed with intelligence for agentic and multimodal tasks. Gemini 3.5 Flash-Lite gemini-3.5-flash-lite minimal $0.30/1M input tokens and $2.50/1M output tokens The fastest, lowest-cost 3.5 model for high-throughput execution. Both models support the 1M token context window, 64k max output tokens, thinking, and the full suite of built-in tools including Computer Use .
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in