When several people talk at once, a transcript can get messy fast.
This is a AI post classified by Jev as Voice AI (a model release), kept by the AI Radar because it carries real work, not commentary.
When several people talk at once, a transcript can get messy fast. Our new Nemotron 3 Diarization model tracks who spoke when, even when voices overlap. It handles up to eight speakers, has 100M parameters, and is now available on @huggingface 🤗
Posted by NVIDIA AI (341.4k followers) 1 h ago · 370 likes · 14.6k views · view the original post on X. Kept by the AI Radar as Voice AI.
More AI work like this
- Google DeepMind's Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS take #1 on all… — @voicearena_ai
- Google发布Gemini 3.8 Flash tts和Flash-lite tts — @mylifcc
- Google has released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. Gemini 3.8 Flash… — @ArtificialAnlys
- Gemini 3.8 Flash TTS & Flash-Lite TTS are live on Google AI Studio and Google Cloud — @thtbee_
- Marking our third audio model release in < 30 days, today we're introducing two new… — @NewsFromGoogle
- GOOGLE 🔥: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are now live on Google AI… — @testingcatalog
- We’re launching Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS ⚡️ — @GoogleAI
- Create and deploy custom audio with our new text-to-speech models: — @GoogleDeepMind
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 44.5k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 16:21 UTC. Full method.