Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

MODEL · ENTITY #7057

Nemotron 3 Diarization

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

MODELS · 1 SOURCE · The Decoder

Nvidia releases Nemotron 3 Diarization, a 100M-parameter real-time speaker diarization model

Nvidia released Nemotron 3 Diarization, a ~100 million-parameter model whose weights are freely available. The model can identify up to eight speakers (including overlapping speech), works on live and recorded audio with configurable buffer sizes, and achieves a 14.72% error rate on VoiceArena's Diarization-Bench, outperforming the prior best system and reducing error versus Streaming Sortformer by about 41% on certain tests.

7.0

MODELS · 1 SOURCE · Hugging Face

NVIDIA releases Nemotron 3 Diarization — open-weight 100M model for real-time multi-speaker diarization

NVIDIA published Nemotron 3 Diarization, an open-weight, 100M-parameter speaker-diarization model that ranks #1 on Voice Arena's Diarization-Bench (14.72% DER). The model supports up to eight anonymous speaker channels, handles overlapping speech in streaming and offline modes, and uses arrival-order speaker caching (AOSC) and a FIFO context buffer; training included public and licensed data, with David AI data reducing compound DER by 0.77 points.

7.0