This model is a fine-tuned version of the wav2vec2-large-robust model on the LibriTTS and VoxPopuli datasets. It is designed for high-quality automatic speech recognition in English and is compatible with the Hugging Face Transformers library for easy local inference.
Wav2vec2 Large Robust Ft Libritts Voxpopuli sits in PulseGate's Voice, TTS & speech category. It focuses on performing high-accuracy English speech-to-text transcription on diverse or noisy audio using a robust fine-tuned model. It is built as an open-source project for developers. Wav2vec2 Large Robust Ft Libritts Voxpopuli is open source under the Open Source license. The product ships for the web and the command line.
jbetker builds and maintains Wav2vec2 Large Robust Ft Libritts Voxpopuli, and the product first shipped in 2021. The GitHub repository has been archived. Key capabilities include Automatic Speech Recognition, fine-tuned on LibriTTS, and robust to Noise.
Latest indexed changes and source events
jbetker/wav2vec2-large-robust-ft-libritts-voxpopuli verified by the PulseGate indexer
⚠ Archived
Other apps tracked under the same category.