Google's New Gemini 3.8 Flash Thinks Harder, But That Thinking Isn't Free
Google has rolled out Gemini 3.8 Flash, replacing 3.7 Flash just weeks after that model's debut. The headline change isn't a price cut or a benchmark flex — it's that the model is designed to "work harder" by running more reasoning steps and calling external tools repeatedly when tackling complex tasks.
Pricing stays the same as 3.7 Flash: $0.75 per million input tokens and $3.75 per million output tokens. But Google itself cautions that the model may burn through more tokens to hit peak performance, particularly at higher "effort" settings, meaning actual costs could climb even though the rate card didn't change. Developers who want predictable, lower token usage can stick with 3.7 Flash instead.
The release has already drawn some early scrutiny from developers testing the tradeoffs in practice.