Kokoro TTS converts text into high-quality speech using AI models. It addresses the need to generate audio from written material such as audiobooks, articles, and other long-form content.
The service provides a selection of voices including American female options Sky, Bella, Nicole, and Sarah, American male options Adam and Michael, British female options Emma and Isabella, and British male options George and Lewis. Users choose between two processing modes: Local VPS, which runs on CPU and is slower, or Hugging Face, which requires zero GPU and processes faster. Text-to-speech computation occurs server-side for improved reliability. The system uses local models; the initial request may require initialization time while later requests complete more quickly.
Input occurs either by uploading TXT or MD files up to 10 MB or by entering text directly with a limit of 500 characters. For larger uploads the tool automatically splits content into chunks and produces separate audio files to preserve quality and keep outputs manageable. The interface includes options to view all voices and to clear input.
It is delivered as a web-based application built for Algoran and described as ready to use. The page indicates the TTS system is running with local models.
In the Text to speech space, Kokoro TTS takes a focused approach. Allowing users to convert large text files into studio-quality speech for audiobooks and content without API fees. It is built as a consumer product for content creators and agencies needing TTS. Kokoro TTS is free to use. Kokoro TTS is available on the web.
Behind Kokoro TTS is Algoran, and it first shipped in 2024. Key capabilities include text-to-speech, batch file processing, and multiple voices.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do