This is a GGUF-quantized version of NVIDIA's Nemotron-3.5 0.6B automatic speech recognition model. It supports streaming transcription in 28 languages using Conformer-Transducer architecture. The model is optimized for local and on-device inference with tools such as transcribe.cpp and is suitable for developers building offline or privacy-focused voice applications.
In the Speech to text space, Nemotron 3.5 Asr Streaming 0.6b takes a focused approach. It focuses on converting streaming speech audio into accurate text across 28 languages without depending on closed cloud transcription services. It is built as an open-source project for developers. The project is open source (MIT). It runs on the web, the command line, and API.
It is developed by handy-computer, and it first shipped in 2026. The project is developed in the open on GitHub with 1.6k stars and 410 commits in the last 90 days. PulseGate's similarity index places it among 8 comparable projects. Among its 4 catalogued features are speech-to-Text, Streaming ASR, and Multilingual Support.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do