FrontierThe story, in brief

NVIDIA Releases NemotronLabs VoiceChat 11B: An Open Full-Duplex Speech-to-Speech Model with ~450 ms Turn-Taking and Live Tool Calling

NVIDIA's 11B speech-to-speech model hits 448ms turn-taking—approaching human conversation speed in open weights.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

A capable open-weight speech model with near-real-time latency and tool integration raises the bar for voice AI capabilities accessible outside closed ecosystems. Practitioners can now evaluate full-duplex voice for production without vendor lock-in.

The key facts

5 to know
  1. NemotronLabs VoiceChat 11B: open full-duplex speech-to-speech model

  2. 448 ms turn-taking latency (near-human conversation threshold)

  3. 11B parameters (efficient size for deployment)

  4. Live tool calling capability

  5. Open weights release

Go to the source

MarkTechPostmarktechpost.com

Publisher excerpt: NVIDIA releases NemotronLabs VoiceChat 11B, an open full-duplex speech-to-speech model with 448 ms latency and live tool calling. The post NVIDIA Releases NemotronLabs VoiceChat 11B: An Open Full-Duplex Speech-to-Speech Model with ~450 ms Turn-Taking and Live Tool Calling appeared first on…
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier