Nemotron Speech Streaming is an NVIDIA-hosted demo that transcribes speech from a microphone or uploaded audio files in real time. It delivers clear, readable text output without delay. The tool showcases NVIDIA's speech recognition models and is available as a Hugging Face Space.
Nemotron Speech Streaming is a Voice, TTS & speech product. It focuses on converting spoken language from live microphone or audio files into accurate text instantly. It is built as an open-source project for developers. The product is available for free. It runs on the web, and it can be self-hosted.
NVIDIA builds and maintains Nemotron Speech Streaming, and the product first shipped in 2024. Key capabilities include Live Transcription, File Upload, and Real-time Output.
Latest indexed changes and source events
nvidia/nemotron-speech-streaming-en-0.6b verified by the PulseGate indexer
Other apps tracked under the same category.