NVIDIA launches open-weight audio-native Nemotron models (2B, 30B)
The open-weight Nemotron models cover transcription, translation, sound recognition, audio Q&A, TTS, and full speech-to-speech. The speech stack is already quantized to GGUF for on-device use via NeMo-Speech.cpp, including Parakeet ASR models and Magpie-TTS Multilingual.
How this story unfolded
3 weeks · 2 reports · 2 community posts · from Jul 22
- Jul 22
- Aug 4
- Aug 6
- Aug 10
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Ante is a coding agent in a single binary that runs offline
- Demo converts static images into editable DrawIO XML via SAM 3
- Lindy reads Slack history and answers with source links
- VicOne releases free cybersecurity extension for NVIDIA Isaac Sim
- Curated collection tracks LLM-based software vulnerability detection