TLDRocket
Sign in

Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

The Verge Stevie Bonifield Covered by 3 sources

Google just launched Gemini 3.8 Flash, a few weeks after 3.7 Flash. It keeps the same starting price, but it may burn more tokens to do the job.

Based on reporting by The Verge, Stevie Bonifield — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

Google has moved fast again: Gemini 3.8 Flash is out only a few weeks after Gemini 3.7 Flash. The pitch is simple enough. This version “works harder” on tougher tasks by taking more reasoning steps and calling tools again and again until it gets where it wants to go.

On paper, the pricing hasn’t changed. Google says 3.8 Flash starts at $0.75 per million input tokens and $3.75 per million output tokens, the same as 3.7 Flash. But the company is also warning that the model may use more tokens to squeeze out better results, especially when users turn up the effort level.

That makes this less like a price cut and more like a tradeoff dressed up as an upgrade. If the model spends longer thinking, the bill can climb even when the sticker price stays put. Google is at least being blunt about that, which is refreshing in a field that usually prefers glossy promises.

Developers who care more about token use than performance can stick with Gemini 3.7 Flash. That option matters, because the new model’s gains may come with a bit of hidden appetite. And in AI, hidden appetite is often where the real cost lives.

My take — AI-written commentary, not fact-checked reporting

This is the kind of product move that should make developers squint. Same starting price, but maybe more tokens, is not exactly a gift from the efficiency gods. Google is basically telling everyone that “better” may also mean “more expensive,” which is the most honest sentence in the whole pitch.

Read more about this at: The Verge

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.