FrontierThe story, in brief

Google launches two benchmark-topping speech generation models

Google's Gemini 3.8 Flash TTS models top every speech-generation benchmark—and cost half as much as competitors.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Two new text-to-speech models from Google set capability benchmarks while offering a cost/quality tradeoff that practitioners will evaluate for production voice workloads. The dual-model release (performance vs. efficiency) signals how frontier labs are now shipping capability ladders rather than single releases.

The key facts

12 to know
  1. Two new models: Gemini 3.8 Flash TTS (best quality) and Gemini 3.8 Flash-Lite TTS (cost/speed optimized)

  2. Both models benchmark-topping in speech generation

  3. Unified API design for side-by-side developer use

  4. Available on Google Cloud Platform

  5. Flash-Lite TTS optimized for inference speed and cost efficiency

  6. Released September 23, 2026

  7. Two new models: Gemini 3.8 Flash TTS (quality-optimized) and Gemini 3.8 Flash-Lite TTS (cost/speed-optimized)

  8. Both models report benchmark-topping performance

  9. Unified API across both models simplifies side-by-side deployment

  10. Available through Google Cloud platform

  11. Flash-Lite TTS targets inference speed and cost efficiency

  12. Flash TTS prioritizes audio quality

Go to the source

SiliconAnglesiliconangle.com

Publisher excerpt: Google LLC today made two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, available through its cloud platform. The algorithms have highly similar application programming interfaces, which makes using them side-by-side relatively simple for developers. Flash-Lite TTS…
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier