NemotronLabs-VoiceChat-11B-mlx Collection Duplex Voice Chat with SSM On-Device • 3 items • Updated 10 days ago • 2
Nemotron Speech Collection Open, state-of-the-art, production‑ready enterprise speech models from the NVIDIA Speech research team for ASR, TTS, Speaker Diarization and S2S • 14 items • Updated 3 days ago • 63
Tiny Series Collection Tiny datasets that empower the foundation of Small Language Model! • 14 items • Updated May 13 • 45
Interactivity Alignment Collection Full-duplex speech models post-trained with reinforcement learning for improved conversational interactivity. • 4 items • Updated about 1 month ago • 6
Stable Audio 3 Extra Collection Contains all checkpoints that are not the standard post-trained checkpoints found in https://huggingface.co/collections/stabilityai/stable-audio-3 • 7 items • Updated May 20 • 12
Nemotron-Post-Training-v3 Collection Collection of datasets used in the post-training phase of Nemotron Nano, Super, and Ultra v3. • 50 items • Updated 3 days ago • 195
SpectroStream: A Versatile Neural Codec for General Audio Paper • 2508.05207 • Published Aug 7, 2025 • 3
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion Paper • 2503.11576 • Published Mar 14, 2025 • 166