Nvidia releases free 100M-parameter model that identifies up to 8 speakers in real time
The Decoder · rss · 2026-09-27
Nvidia has released Nemotron 3 Diarization, a free AI model that performs speaker diarization—identifying who is speaking at any given moment in a conversation.
The model is only about 100M parameters yet can distinguish up to eight speakers in real time, making it a lightweight, practical building block for transcription, captioning, and voice-analytics applications.
More from Models
- rasbt and marlene_zw break down Claude watermarks, reasoning models in new TechTalk — marlene_zw · 2026-09-27
- Opus 5.5 praised as remarkably efficient: top-tier quality at surprisingly good rates — kimmonismus · 2026-09-27
- Local AI community urges Qwen to bring back a 35B-class MoE for low-VRAM GPUs — julianharris · 2026-09-27
- Grok 4.7 lifts Terminal-Bench 4.0 from 20.3% to 38%, but burns 125% more output tokens — dl_weekly · 2026-09-27
- TypeSafe's Jev: A Decision-Only Model That Returns Typed Choices Instead of Generated Text — Rahulstark2 · 2026-09-27
- Codex team hints point to speed — GPT-6 Astra on Cerebras rumored — haider1 · 2026-09-27