Voice, TTS & speech Tools & Software

Live Voice, TTS & speech listings on PulseGate.

Voice, TTS & speech is part of AI on PulseGate. PulseGate tracks 1,371 Voice, TTS & speech products — 13 indexed in the past week, most recently Musicgen Large.

Musicgen Large
huggingface.co
Voice, TTS & speech · 12h ago

musicgen-large is a large text-to-music generation model developed by Meta.

Voice, TTS & speech
12h ago
Speaker Lite logo
Speaker Lite
speaker-lite.merkulov.design
Voice, TTS & speech · 13h ago

Speaker Lite converts WordPress page content into human-like speech using Google Cloud TTS.

Voice, TTS & speech
13h ago
Nb Wav2vec2 1b Nynorsk
huggingface.co
Voice, TTS & speech · 1d ago

nb-wav2vec2-1b-nynorsk is a Norwegian (Nynorsk) speech recognition model based on wav2vec2.

Voice, TTS & speech
1d ago
LocalVoice logo
LocalVoice
hajeklabs.com
Voice, TTS & speech · 3d ago

LocalVoice is an offline AI text-to-speech and voice cloning app for macOS.

Voice, TTS & speech
3d ago
Ainnate Text To Speech logo
Ainnate Text To Speech
tts.ainnate.com
Voice, TTS & speech · 3d ago

Ainnate Text To Speech provides advanced AI voice synthesis technology for generating natural-sounding speech from text.

Voice, TTS & speech
3d ago
kokoro-cli
pypi.org
Voice, TTS & speech · 4d ago

kokoro-cli is an offline-first text-to-speech CLI tool and localhost service powered by Kokoro.

Voice, TTS & speech
4d ago
MOSS TTS — huggingface.co/OpenMOSS-Team/MOSS-TTS
huggingface.co
Voice, TTS & speech · 4d ago

MOSS-TTS is an open-source text-to-speech model developed by the OpenMOSS team.

Voice, TTS & speech
4d ago
Nemotron 3.5 Asr Streaming 0.6b
huggingface.co
Voice, TTS & speech · 4d ago

Nemotron-3.5-ASR-Streaming is an open-source multilingual automatic speech recognition model in GGUF format.

Voice, TTS & speech
4d ago
Voxtral Mini 4B Realtime 2602
huggingface.co
Voice, TTS & speech · 5d ago

Voxtral-Mini-4B-Realtime-2602-gguf is a quantized GGUF model for real-time multilingual speech-to-text and audio understanding.

Voice, TTS & speech
5d ago
VieNeu TTS V3 Turbo
huggingface.co
Voice, TTS & speech · 5d ago

VieNeu-TTS-v3-Turbo is a Vietnamese text-to-speech model supporting voice cloning, emotion control, and high-fidelity 48kHz audio.

Voice, TTS & speech
5d ago
Canary 1b
huggingface.co
Voice, TTS & speech · 5d ago

Quantized GGUF version of the Canary 1B multilingual speech transcription and translation model.

Voice, TTS & speech
5d ago
Wav2vec2 Large Xlsr 53 Finnish
huggingface.co
Voice, TTS & speech · 5d ago

Finnish speech recognition model based on XLS-R wav2vec2.

Voice, TTS & speech
5d ago
Qwen3 TTS 12Hz 0.6B CustomVoice
huggingface.co
Voice, TTS & speech · 5d ago

Qwen3-TTS-12Hz-0.6B-CustomVoice is an open-source text-to-speech model supporting custom voice cloning across multiple languages.

Voice, TTS & speech
5d ago
ovos-skill-andersen-tales
pypi.org
Voice, TTS & speech · 7d ago

ovos-skill-andersen-tales is a provider skill that reads Hans Christian Andersen fairy tales aloud for voice assistants.

Voice, TTS & speech
7d ago
ovos-skill-grimm-tales
pypi.org
Voice, TTS & speech · 7d ago

ovos-skill-grimm-tales is a provider skill that reads Brothers Grimm fairy tales aloud for voice assistants.

Voice, TTS & speech
7d ago
ovos-skill-ovosblog
pypi.org
Voice, TTS & speech · 7d ago

ovos-skill-ovosblog is a provider skill that reads the OpenVoiceOS blog aloud for voice assistants.

Voice, TTS & speech
7d ago
ovos-skill-arxiv-papers
pypi.org
Voice, TTS & speech · 7d ago

ovos-skill-arxiv-papers is a provider skill that reads arXiv paper abstracts aloud for OpenVoiceOS voice assistants.

Voice, TTS & speech
7d ago
Whisper Large
huggingface.co
Voice, TTS & speech · 7d ago

GGUF quantized version of OpenAI's Whisper Large V3 for speech-to-text transcription.

Voice, TTS & speech
7d ago
Wav2vec2 Conformer Rope Large 960h Ft
huggingface.co
Voice, TTS & speech · 7d ago

Large wav2vec2 Conformer model with rotary embeddings fine-tuned for speech recognition.

Voice, TTS & speech
7d ago
Faster Whisper Small.en
huggingface.co
Voice, TTS & speech · 7d ago

Converted Whisper small.en model optimized for CTranslate2 inference.

Voice, TTS & speech
7d ago
Musicgen Small
huggingface.co
Voice, TTS & speech · 7d ago

musicgen-small is a small MusicGen model by Meta for text-to-music generation.

Voice, TTS & speech
7d ago
Faster Whisper Base.en
huggingface.co
Voice, TTS & speech · 7d ago

faster-whisper-base.en is a CTranslate2-converted Whisper base model for fast English speech recognition.

Voice, TTS & speech
7d ago
MOSS TTS — huggingface.co/OpenMOSS-Team/MOSS-TTS-v1.5
huggingface.co
Voice, TTS & speech · 7d ago

MOSS-TTS-v1.5 is an open-source text-to-speech model developed by the OpenMOSS team.

Voice, TTS & speech
7d ago
Parakeetkit Pro
huggingface.co
Voice, TTS & speech · 7d ago

On-device automatic speech recognition models for WhisperKit and Argmax Pro SDK.

Voice, TTS & speech
7d ago
Medasr
huggingface.co
Voice, TTS & speech · 7d ago

A medical automatic speech recognition model developed by Google for radiology and clinical domains.

Voice, TTS & speech
7d ago
Speaker Diarization — huggingface.co/pyannote/speaker-diarization-3.0
huggingface.co
Voice, TTS & speech · 7d ago

speaker-diarization-3.0 is a pipeline for determining who spoke when in an audio recording.

Voice, TTS & speech
7d ago
Mms Lid 256
huggingface.co
Voice, TTS & speech · 7d ago

Facebook's Massively Multilingual Speech model for language identification.

Voice, TTS & speech
7d ago
Wav2vec2 Lv 60 Espeak Cv Ft
huggingface.co
Voice, TTS & speech · 7d ago

wav2vec2-lv-60-espeak-cv-ft is an open-source speech recognition model for low-resource languages and phonetic transcription.

Voice, TTS & speech
7d ago
Wav2vec2 Large Xlsr Catala
huggingface.co
Voice, TTS & speech · 7d ago

wav2vec2-large-xlsr-catala is a speech recognition model fine-tuned for the Catalan language.

Voice, TTS & speech
7d ago
Faster Whisper Medium
huggingface.co
Voice, TTS & speech · 7d ago

faster-whisper-medium is an optimized Whisper medium model for fast speech transcription using CTranslate2.

Voice, TTS & speech
7d ago
Whisper Bemba Stt
huggingface.co
Voice, TTS & speech · 7d ago

whisper-bemba-stt is a fine-tuned Whisper model for automatic speech recognition in the Bemba language.

Voice, TTS & speech
7d ago
Faster Whisper Large V3 Turbo Ct2
huggingface.co
Voice, TTS & speech · 7d ago

Whisper large-v3-turbo model converted for CTranslate2.

Voice, TTS & speech
7d ago
PersonaPlex 7B
huggingface.co
Voice, TTS & speech · 7d ago

PersonaPlex-7B-MLX-4bit is a quantized speech-to-speech model optimized for Apple Silicon.

Voice, TTS & speech
7d ago
tts-daemon
pypi.org
Voice, TTS & speech · 7d ago

tts-daemon is a local HTTP/WebSocket Text-to-Speech gateway with pluggable providers, starting with Piper.

Voice, TTS & speech
7d ago
Qwen3 TTS Tokenizer 12Hz
huggingface.co
Voice, TTS & speech · 7d ago

Tokenizer for Qwen3-TTS enabling extreme bitrate reduction and low-latency speech processing.

Voice, TTS & speech
7d ago
Qwen3 TTS
huggingface.co
Voice, TTS & speech · 7d ago

GGUF quantized versions of Qwen3-TTS text-to-speech models for local inference.

Voice, TTS & speech
7d ago
Parakeet Tdt 0.6b
huggingface.co
Voice, TTS & speech · 7d ago

GGUF quantized version of NVIDIA's Parakeet TDT 0.6B automatic speech recognition model.

Voice, TTS & speech
7d ago
Diar Streaming Sortformer 4spk
huggingface.co
Voice, TTS & speech · 7d ago

NVIDIA's streaming speaker diarization model using Sortformer for up to 4 speakers.

Voice, TTS & speech
7d ago
Wav2vec2 Large Robust Ft Libritts Voxpopuli
huggingface.co
Voice, TTS & speech · 7d ago

wav2vec2-large-robust-ft-libritts-voxpopuli is a fine-tuned speech recognition model for English audio.

Voice, TTS & speech
7d ago
Faster Distil Whisper Medium.en
huggingface.co
Voice, TTS & speech · 7d ago

faster-distil-whisper-medium.en is a CTranslate2-optimized English speech-to-text model based on Distil-Whisper.

Voice, TTS & speech
7d ago
Granite 4.0 1b Speech
huggingface.co
Voice, TTS & speech · 7d ago

granite-4.0-1b-speech is an IBM Granite model for automatic speech recognition tasks.

Voice, TTS & speech
7d ago
Qwen3 ASR 1.7B
huggingface.co
Voice, TTS & speech · 7d ago

Qwen3-ASR-1.7B is an open-source speech recognition model provided in GGUF format for local inference.

Voice, TTS & speech
7d ago
Voxtral Small 24B 2507
huggingface.co
Voice, TTS & speech · 7d ago

Voxtral-Small-24B-2507-gguf provides GGUF quantized versions of an open audio-language model for speech recognition and transcription.

Voice, TTS & speech
7d ago
Nemotron Speech Streaming En 0.6b
huggingface.co
Voice, TTS & speech · 7d ago

nemotron-speech-streaming-en-0.6b is an English automatic speech recognition model from NVIDIA.

Voice, TTS & speech
7d ago
Mms Lid 126
huggingface.co
Voice, TTS & speech · 7d ago

facebook/mms-lid-126 is a massively multilingual speech model for language identification supporting 126 languages.

Voice, TTS & speech
7d ago
Whisper Medium
huggingface.co
Voice, TTS & speech · 7d ago

GGUF quantized version of OpenAI's Whisper medium model for speech recognition.

Voice, TTS & speech
7d ago
Speaker Diarization — huggingface.co/pyannote/speaker-diarization-3.1
huggingface.co
Voice, TTS & speech · 8d ago

Speaker diarization pipeline that identifies who spoke when in audio.

Voice, TTS & speech
8d ago
Kazakh Whisper Large V3 Turbo
huggingface.co
Voice, TTS & speech · 8d ago

A fine-tuned Whisper model for Kazakh speech recognition.

Voice, TTS & speech
8d ago
Qwen3 ASR 0.6B
huggingface.co
Voice, TTS & speech · 9d ago

Qwen3-ASR-0.6B-gguf is a quantized automatic speech recognition model in GGUF format.

Voice, TTS & speech
9d ago
S2t Small Librispeech Asr
huggingface.co
Voice, TTS & speech · 9d ago

S2T Small Librispeech ASR is a speech-to-text model trained on the LibriSpeech dataset for automatic speech recognition in English.

Voice, TTS & speech
9d ago
Showing 150 of 1,371
Explore results. Clear filters to return to Top 500.