Qwen3-TTS-12Hz-0.6B-Base is an open-source text-to-speech foundation model from the Qwen team. It supports English, Chinese, Japanese, Korean and several other languages and can be used for speech synthesis and voice cloning tasks. The model is distributed on Hugging Face and is compatible with standard inference libraries.
In the Text to speech space, Qwen3 TTS 12Hz 0.6B Base takes a focused approach. It focuses on generating natural-sounding speech audio from text in multiple languages using a small open model. Qwen3 TTS 12Hz 0.6B Base is an open-source project aimed at developers. The project is open source (Apache-2.0). It ships for the web, the command line, and API.
Qwen builds and maintains Qwen3 TTS 12Hz 0.6B Base, and it first shipped in 2026. Development happens publicly on GitHub with 12.5k stars. PulseGate's similarity index places it among 8 comparable projects. Among its 3 catalogued features are text-to-Speech, Voice Cloning, and multilingual.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do