Qwen3-ASR-0.6B-hf is an automatic speech recognition model from Alibaba's Qwen team. It supports audio input alongside text and is distributed in Hugging Face format with a chat template that handles audio tokens. The model is suitable for integration into voice-enabled applications and can be run locally or via inference providers.
In the Speech to text space, Qwen3 ASR 0.6B Hf takes a focused approach. It focuses on converting spoken audio into text with a compact multilingual model. It is built as an open-source project for AI developers. The project is open source (Open Source). It runs on the web, the command line, and API.
It is developed by Qwen. PulseGate's similarity index places it among 13 comparable projects. Among its 3 catalogued features are speech recognition, multimodal support, and chat template.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do