facebook/wav2vec2-lv-60-espeak-cv-ft is a fine-tuned wav2vec 2.0 model for automatic speech recognition and phonetic transcription. Trained on 60 languages from the Common Voice and espeak datasets, it is published as open weights on Hugging Face. It is intended for developers building multilingual speech applications using the Transformers library.
In the Speech to text space, Wav2vec2 Lv 60 Espeak Cv Ft takes a focused approach. It focuses on performing accurate speech-to-text and phonetic transcription for diverse and low-resource languages using an open model. Wav2vec2 Lv 60 Espeak Cv Ft is an open-source project aimed at developers. The project is open source (MIT). Wav2vec2 Lv 60 Espeak Cv Ft is available on the web and API.
Meta builds and maintains Wav2vec2 Lv 60 Espeak Cv Ft, and it first shipped in 2017. The GitHub repository has been archived. Key capabilities include Automatic Speech Recognition, Phonetic Transcription, and Low-resource Languages. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do