NemotronLabs-VoiceChat-11B-mlx Collection Duplex Voice Chat with SSM On-Device • 3 items • Updated Aug 5 • 2
Nemotron Speech Collection Open, state-of-the-art, production‑ready enterprise speech models from the NVIDIA Speech research team for ASR, TTS, Speaker Diarization and S2S • 13 items • Updated 21 days ago • 69
Tiny Series Collection Tiny datasets that empower the foundation of Small Language Model! • 14 items • Updated May 13 • 45
Interactivity Alignment Collection Full-duplex speech models post-trained with reinforcement learning for improved conversational interactivity. • 4 items • Updated Jul 15 • 6
Stable Audio 3 Extra Collection Contains all checkpoints that are not the standard post-trained checkpoints found in https://huggingface.co/collections/stabilityai/stable-audio-3 • 7 items • Updated May 20 • 12
Nemotron-Post-Training-v3 Collection Collection of datasets used in the post-training phase of Nemotron Nano, Super, and Ultra v3. • 50 items • Updated about 1 month ago • 200
SpectroStream: A Versatile Neural Codec for General Audio Paper • 2508.05207 • Published Aug 7, 2025 • 4
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion Paper • 2503.11576 • Published Mar 14, 2025 • 174