Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

Google has launched Gemini 3.8 Flash, a new iteration of its large language model that arrives shortly after the release of the…

By Vane September 2, 2026 1 min read
Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

Google has launched Gemini 3.8 Flash, a new iteration of its large language model that arrives shortly after the release of the 3.7 version. The company states this update works harder by executing more reasoning steps and calling tools iteratively to handle complex tasks. While the introductory pricing matches the previous model at $0.75 per million input tokens and $3.75 per million output tokens, Google warns that users could face higher costs. The new model might consume more tokens to maximise performance, particularly at higher effort levels. Developers wishing to minimise token usage can continue using Gemini 3.7 Flash instead.

This shift highlights a trade-off between performance and efficiency that developers must consider when selecting models for production. Choosing a model that works harder does not always mean better value if the increased token consumption outweighs the quality gains. Teams need to evaluate whether the extra reasoning steps justify the potential rise in operational expenses.

* Pricing remains identical to Gemini 3.7 Flash at launch
* Higher effort levels may increase token consumption
* Older models remain available for cost-sensitive projects

Scroll to Top