ToolsThe story, in brief

Introducing next-generation audio models in the API

OpenAI just unlocked style-based voice customization. Developers can now tell TTS models to speak like 'a sympathetic customer service agent'—turning generic audio into persona-driven interactions.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI's next-gen audio APIs move text-to-speech from one-size-fits-all to instruction-based voice personas, enabling founders to ship voice agents with custom tone and personality without custom training. This is a direct competitive move in the emerging voice-agent stack.

The key facts

4 to know
  1. Text-to-speech model now accepts style instructions (e.g., 'sympathetic customer service agent')

  2. First time developers can customize voice behavior via API prompt

  3. Voice agent customization without retraining

  4. Announced March 20, 2025 via OpenAI official channel

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: For the first time, developers can also instruct the text-to-speech model to speak in a specific way—for example, “talk like a sympathetic customer service agent”—unlocking a new level of customization for voice agents.
Read original report
Back to today's editionMore tools news

Keep reading

Related stories

More from Tools