FrontierSeptember 2, 2026via The Verge AI

Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

Why it matters

A new capability tier (extended reasoning, iterative tool use) arrives with same-tier pricing but hidden cost structure: users pay per token, not per capability. Practitioners need to benchmark token expansion against performance gains before upgrading.

Key signals

  • Gemini 3.8 Flash released weeks after 3.7 Flash
  • Claims: performs more reasoning steps, calls tools iteratively on complex tasks
  • Pricing: $0.75 per million input tokens, $3.75 per million output tokens (same as 3.7 Flash)
  • Warning: model may use more tokens at higher effort levels, increasing effective cost
  • Users can opt to keep using 3.7 Flash to control token spend
  • Capability upgrade (reasoning depth, tool iteration) is the story; pricing structure creates practitioner decision point
  • Gemini 3.8 Flash launched weeks after 3.7 Flash
  • Performs more reasoning steps and iterative tool calls on complex tasks
  • Introductory pricing: $0.75 per million input tokens, $3.75 per million output tokens (same as 3.7)
  • Google warns model may consume more tokens at higher effort levels
  • Developers can revert to 3.7 Flash to minimize token usage
  • Release timing suggests rapid iteration cycle in frontier model development

The hook

Google's Gemini 3.8 Flash 'works harder'—but your token bill might not like it.

Google launched Gemini 3.8 Flash, arriving just a few weeks after its predecessor. The company claims the new model "works harder" than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and "calling tools iteratively." It has the same introductory pricing as 3.7 Flash, $0.75 per m

The week's key stories, every Friday.

ONE BRIEFING · EVERY FRIDAY · FREE

Free. Unsubscribe anytime.