nemotron-speech-streaming-en-0.6b is a compact English automatic speech recognition model developed by NVIDIA using the NeMo framework. It supports streaming inference for real-time transcription and is suitable for integration into applications requiring low-latency speech recognition. The model is publicly available on Hugging Face for local or self-hosted deployment.
Nemotron Speech Streaming En 0.6b is a Voice, TTS & speech product. It focuses on converting English speech to text in real-time streaming scenarios without using proprietary cloud services. It is built as an open-source project for developers. Nemotron Speech Streaming En 0.6b is open source under the Open Source license. The product ships for the command line and API.
Behind Nemotron Speech Streaming En 0.6b is NVIDIA, based in the United States, and the product first shipped in 2025. Development happens publicly on GitHub with 42 stars. Key capabilities include Speech Recognition, Streaming ASR, and NVIDIA NeMo.
Latest indexed changes and source events
nvidia/nemotron-speech-streaming-en-0.6b verified by the PulseGate indexer
Other apps tracked under the same category.