{"rewrite":{"id":"r_66d04774af64268c1dae6892","clusterId":"c_f653f296e1eecad039b9b2a1","slug":"nvidia-releases-nemotron-3-diarization-as-an-open-model","model":"deepseek-v4-1-flash","headline":"NVIDIA Releases Nemotron 3 Diarization As An Open Model","summary":"NVIDIA released Nemotron 3 Diarization, a speech recognition model that identifies speakers while transcribing in real time. It handles up to eight speakers and labels them in order of appearance as Speaker 1, Speaker 2, and so on, without registering voices in advance. Training used public datasets and data licensed from David AI. The model is open, under the OpenMDW-1.1 license, and available on Hugging Face.","whyItMatters":"NVIDIA is competing in real-time transcription with an open license and a speaker-ID benchmark claim, which puts an eight-speaker model in reach of anyone building meeting, captioning, or archive tools without a hosted API.","webCardHtml":"\u003cp\u003eThe model splits speaker identification from transcription instead of running both as one pass, and it does not need each voice registered beforehand. It labels appearing speakers in sequence, so a recording with unknown participants can be divided by voice as it is transcribed.\u003c/p\u003e\u003cp\u003eNVIDIA says it trained the model on public datasets plus datasets licensed from David AI. The release includes a Hugging Face Space demo and the weights under the OpenMDW-1.1 license, with a YouTube demonstration showing speaker changes recognized during transcription.\u003c/p\u003e","blueskyPost":"NVIDIA released Nemotron 3 Diarization as an open model under OpenMDW-1.1. It identifies up to 8 speakers while transcribing in real time and labels them in order of appearance, with no voice registration first.","twitterPost":"NVIDIA released Nemotron 3 Diarization as an open model under OpenMDW-1.1. It identifies up to 8 speakers while transcribing in real time and labels them in order of appearance, with no voice registration first.","threadsPost":null,"newsletterBlurb":"NVIDIA has released Nemotron 3 Diarization, a speech model that identifies up to eight speakers while transcribing in real time. It splits speaker identification from transcription and assigns labels in order of appearance rather than requiring registered voices. The model ships as an open release under the OpenMDW-1.1 license with a Hugging Face demo.","attributionJson":"[{\"source\":\"GIGAZINE\",\"url\":\"https://gigazine.net/news/20260924-nvidia-nemotron-3-diarization/\",\"title\":\"NVIDIA Nemotron 3 Diarization, an AI model that can identify up to 8 speakers while transcribing in real time, is released as an open model\"}]","lintFlagsJson":null,"lintHits":0,"costUsd":0,"inputTokens":4736,"outputTokens":574,"status":"published","repairAttempts":0,"nextRepairAt":null,"factsAttemptedAt":1791214634,"createdAt":"2026-10-05T15:32:55.000Z","publishedAt":"2026-10-05T15:36:07.000Z","updatedAt":"2026-10-05T15:36:07.000Z"},"cluster":{"id":"c_f653f296e1eecad039b9b2a1","canonicalTitle":"最大8人の話者を識別しつつリアルタイム文字起こしできるAIモデル「NVIDIA Nemotron 3 Diarization」がオープンモデルとして公開される","representativeArticleId":"a_43bc520e878300640ee8ea10","sourceCount":1,"writtenSourceCount":1,"writeAttempts":0,"isSolo":true,"entitiesJson":"{\"anime_titles\":[],\"manga_titles\":[],\"work_titles\":[\"NVIDIA Nemotron 3 Diarization\"],\"studios\":[\"NVIDIA\"],\"people\":[],\"type\":\"announcement\",\"domain\":\"other\",\"is_roundup\":false}","contentType":"news","status":"published","firstSeenAt":"2026-09-24T09:15:00.000Z","lastSeenAt":"2026-09-24T09:15:00.000Z","updatedAt":"2026-10-05T15:36:06.000Z"},"attribution":[{"source":"GIGAZINE","url":"https://gigazine.net/news/20260924-nvidia-nemotron-3-diarization/","title":"最大8人の話者を識別しつつリアルタイム文字起こしできるAIモデル「NVIDIA Nemotron 3 Diarization」がオープンモデルとして公開される"}],"entities":{"anime_titles":[],"manga_titles":[],"work_titles":["NVIDIA Nemotron 3 Diarization"],"studios":["NVIDIA"],"people":[],"type":"announcement","domain":"other","is_roundup":false},"keyFacts":null}
