Qwen3-ASR-0.6B is a lightweight automatic speech recognition model released by Alibaba's Qwen team. It processes audio inputs to generate text transcriptions and is optimized for efficiency while maintaining strong performance across languages. Available on Hugging Face, it supports integration via pip, Docker, and cloud inference endpoints for developers building voice-enabled applications.
In the Voice, TTS & speech space, Qwen3 ASR 0.6B takes a focused approach. It focuses on converting spoken audio into accurate text transcripts for developers building voice applications. Qwen3 ASR 0.6B is an open-source project aimed at developers. The project is open source (BSD-3-Clause). The product ships for the web, the command line, and API.
It is developed by Alibaba (China), and the product first shipped in 2022. The project is developed in the open on GitHub with 24.5k stars and 73 commits in the last 90 days. PulseGate's similarity index places it among 16 comparable tools. Among its 4 catalogued features are Speech Recognition, Audio Processing, and 0.6B Parameters. It exposes integrations via a public API.
Latest indexed changes and source events
Qwen/Qwen3-ASR-0.6B verified by the PulseGate indexer
Other apps tracked under the same category.