FrontierSeptember 15, 2026via MarkTechPost

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

Why it matters

Google ships a production-grade voice model with extended reasoning and tool calling, positioning itself for the agent-native era. Practitioners can deploy multimodal, multilingual voice agents immediately; enthusiasts see another major lab capability release in what's becoming a weekly cadence of frontier model drops.

Key signals

  • Gemini 3.8 Live and 3.8 Live Extended Thinking released to production
  • Extended Thinking ranks #1 on Artificial Analysis' Speech to Speech Quality Index with 82.6
  • Scores 97.7% on Big Bench Audio benchmark
  • Supports 97 languages with mid-conversation switching
  • Tool and API execution in background during conversation
  • Live visual input processing
  • Pricing: $0.005/min for audio input
  • Available in Gemini API and Google AI Studio
  • Audio output watermarked with Google DeepMind's SynthID

The hook

Google's Gemini 3.8 Live Extended Thinking hits 82.6 on speech-quality benchmarks — and it's in production today at $0.005/min.

Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to date. The models execute tools and API calls in the background while the conversation keeps flowing, process live visual inputs, and switch between 97 languages mid conversation. Exte

The week's key stories, every Friday.

ONE BRIEFING · EVERY FRIDAY · FREE

Free. Unsubscribe anytime.