# Nvidia releases open model that labels up to eight speakers in live audio

_Published Sunday, September 27, 2026 at 8:19 AM EDT · AI · Latest · Tier 2 — Notable_

Nvidia has released Nemotron 3 Diarization, a freely available model that identifies when each of up to eight people speaks in recorded or live audio. Paired with speech recognition software, it can produce transcripts with speaker labels, though those labels identify speakers only as anonymous entries such as "speaker_2."

The model's audio buffer can be set between 30.4 and 0.32 seconds, with shorter buffers generally reducing accuracy. At a 1.04-second buffer, it cut the error rate of its predecessor by an average of 41 percent across eight test scenarios. More speakers, heavy background noise and reverb raise error rates.

## Sources

- [the-decoder.com](https://the-decoder.com/nvidia-drops-a-free-100m-parameter-model-that-identifies-up-to-eight-speakers-in-real-time/)

---
Canonical: https://techandbusiness.org/newswire/0bSk3xEojovOKLM5pi37yu
Published: 2026-09-27T12:19:50.506Z
Story chronology: 2026-09-27T11:01:13.000Z
Retrieved: 2026-09-27T15:11:57.686Z
Publisher: Tech & Business (techandbusiness.org)
