NVIDIA Releases Nemotron 3 Diarization As An Open Model
NVIDIA released Nemotron 3 Diarization, a speech recognition model that identifies speakers while transcribing in real time. It handles up to eight speakers and labels them in order of appearance as Speaker 1, Speaker 2, and so on, without registering voices in advance. Training used public datasets and data licensed from David AI. The model is open, under the OpenMDW-1.1 license, and available on Hugging Face.