Nvidia has announced the release of Nemotron 3 Diarization, a new AI model designed to identify which speaker is talking in real time during conversations. The model, which boasts 100 million parameters, can differentiate between up to eight speakers, making it a significant advancement in audio processing technology.

This release is particularly noteworthy as it is being offered for free, allowing a wider range of developers and researchers to leverage its capabilities. The introduction of this model could enhance applications in various fields, including customer service, transcription services, and collaborative environments where multiple speakers are present.


Compiled automatically by the Tech AI Newsdesk from public AI-news sources and summarised in our own words.