Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

Google has released Gemini 3.8 Flash, a new AI model that performs more reasoning steps and iterative tool usage compared to its predecessor. While the base token price remains the same, the model's increased performance may lead to higher overall costs for developers.
Why it matters
This release marks a shift in AI development toward 'agentic' models that prioritize complex reasoning and task execution over simple text generation.
Google launched Gemini 3.8 Flash, arriving just a few weeks after its predecessor. The company claims the new model “works harder” than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and “calling tools iteratively.” It has the same introductory pricing as 3.7 Flash, $0.75 per million input tokens and $3.75 per million output tokens, but could still end up costing users more. Google warns that “the model might use more tokens to maximize performance, especially at higher effort levels.” Developers can keep using Gemini 3.7 Flash if they want to minimize token usage.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in