Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

https://platform.theverge.com/wp-content/uploads/sites/2/2025/02/STK255_Google_Gemini_B_474198.jpg?quality=90&strip=all&crop=0%2C10.732984293194%2C100%2C78.534031413613&w=1200

Google launched Gemini 3.8 Flash, arriving just a few weeks after its predecessor. The company claims the new model “works harder” than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and “calling tools iteratively.” It has the same introductory pricing as 3.7 Flash, $0.75 per million input tokens and $3.75 per million output tokens, but could still end up costing users more. Google warns that “the model might use more tokens to maximize performance, especially at higher effort levels.” Developers can keep using Gemini 3.7 Flash if they want to minimize token usage.

Gemini 3.8 Flash’s launch was followed by some early impressions online. Artificial Analysishighlighted the model’s pricing, saying Gemini 3.8 Flash is “the cheapest we’ve measured at this level of intelligence. This is up ~40% from Gemini 3.7 Flash despite unchanged per-token pricing, driven by a 30% increase in output tokens per...

Copyright of this story solely belongs to theverge.com. To see the full text click HERE

Read more