VITS Models is a text-to-speech application supporting Chinese and Japanese input. Users select different speakers and languages to generate high-quality audio output. It is hosted as a Hugging Face Space using open-source VITS models.
Vits Models is a Voice, TTS & speech project. It focuses on generating natural speech audio from text in East Asian languages. Vits Models is an open-source project aimed at content creators and language learners. Vits Models costs nothing to use. It runs on the web, and it can be self-hosted.
It is developed by zomehwh, and it first shipped in 2023. Among its 3 catalogued features are Voice Synthesis, multi-speaker, and chinese and Japanese Support.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do