Google's Gemini 3.8 Flash 'works harder' but may burn more tokens at same pricing

The Verge AI · rss · 2026-09-03

Just weeks after Gemini 3.7 Flash, Google launched Gemini 3.8 Flash, claiming it "works harder" by running more reasoning steps and calling tools iteratively on complex tasks. Intro pricing matches 3.7 Flash: $0.75/M input and $3.75/M output tokens. But Google warns the model may use more tokens to maximize performance, especially at higher effort levels, so costs could climb. Developers wanting to minimize token usage can stay on 3.7 Flash. Early reactions to the launch have been mixed.

Original post →

More from Infra

Infra channel →