FrontierThe story, in brief

**Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization**

NVIDIA open-sources Nemotron 3 diarization — speaker identification at real-time speed now available to builders.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Speaker diarization (who said what, when) is table-stakes for voice AI products. NVIDIA's open-weight release lowers the barrier to multi-speaker audio understanding for practitioners building voice agents, meeting transcription, and conversational AI.

The key facts

9 to know
  1. NVIDIA Nemotron 3 diarization model open-sourced

  2. Real-time multi-speaker identification capability

  3. Open-weight release on Hugging Face

  4. Addresses speaker tracking in audio streams

  5. Published September 23, 2026

  6. NVIDIA Nemotron 3 diarization — open-weight release

  7. Real-time, multi-speaker capability

  8. Published Sept 23, 2026 on Hugging Face

  9. Reduces dependency on proprietary speaker-diarization APIs for voice AI builders

Go to the source

Hugging Face Bloghuggingface.co

Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier