Google’s Gemini 3.8 Flash Works Harder
Google's new Gemini 3.8 Flash model performs more reasoning steps but may increase token usage.

The update
Google has launched Gemini 3.8 Flash, arriving just weeks after its predecessor. The company claims the new model “works harder” than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and “calling tools iteratively.” It maintains the same introductory pricing as 3.7 Flash ($0.75 per million input tokens and $3.75 per million output tokens), but Google warns that “the model might use more tokens to maximize performance, especially at higher effort levels.”
Why it matters
This rapid release cycle—Google’s third Flash model in six weeks—suggests an industry-wide focus on balancing cost and performance. The model shows significant improvements in software engineering and autonomous AI agent benchmarks, outperforming competitors like Anthropic’s Fable 5 on certain tests. Early impressions indicate it offers “Opus 5 coding quality but at a fraction of the cost,” according to one AI expert.
What to watch
Will developers stick with 3.7 Flash to minimize token usage? How will pricing evolve after the introductory period? The model’s safeguards against misuse in cybersecurity and other sensitive domains may also impact adoption.
Sources
- theverge.com — pricing details and performance claims
- arstechnica.com — model variations and benchmark performance
How did this story land?
Choose one reaction. Choosing it again leaves it selected.
