An automatic speech recognition (ASR) model developed by Microsoft that converts audio input into text. It is designed to output structured JSON and includes a specialized chat template for handling audio tokens. The model is published on Hugging Face and can be used with the transformers library or compatible inference frameworks.
VibeVoice ASR HF is an Other AI product. It focuses on converting spoken audio into accurate, structured text transcriptions. VibeVoice ASR HF is an open-source project aimed at developers building voice applications. The project is open source (MIT). VibeVoice ASR HF is available on the web and API.
It is developed by Microsoft (United States), and the product first shipped in 2025. The project is developed in the open on GitHub with 50.1k stars and 2 commits in the last 90 days. Among its 3 catalogued features are Automatic Speech Recognition, JSON Output, and Audio Transcription.
Latest indexed changes and source events
microsoft/VibeVoice-ASR-HF verified by the PulseGate indexer
Other apps tracked under the same category.