Google launched Gemini 3.8 Flash just a few weeks after releasing its predecessor, Gemini 3.7 Flash. According to the company, the new model “works harder” by performing additional reasoning steps on complex tasks and calling tools iteratively. It carries the same introductory pricing as 3.7 Flash, listed at 3.75 per million output tokens.
Despite the matching price, Google cautions that costs could rise for some users because the model may consume more tokens to maximize performance, particularly at higher effort levels. Developers who want to limit token usage can continue using Gemini 3.7 Flash.
Why it matters
The pricing note highlights a tradeoff between performance and cost: even with unchanged per-token rates, increased token consumption can lead to higher bills. Google’s decision to keep the older model available gives developers a way to manage that tradeoff.
Who should care
Developers and teams building on the Gemini API should weigh the potential for higher spending against any performance gains, and decide whether to adopt 3.8 Flash or stay on 3.7 Flash.