wav2vec2-large-xlsr-53-mongolian is a fine-tuned version of the XLS-R wav2vec2 model trained on Mongolian speech data from the Common Voice dataset. It supports automatic speech recognition tasks and is available on Hugging Face for use with the Transformers library. The model was created by Anton Lozhkov to improve speech-to-text performance for low-resource languages like Mongolian.
Wav2vec2 Large Xlsr 53 Mongolian is a Speech to text project. It focuses on enabling accurate automatic speech recognition for Mongolian language audio. Wav2vec2 Large Xlsr 53 Mongolian is an open-source project aimed at developers. The project is open source (Open Source). Wav2vec2 Large Xlsr 53 Mongolian is available on the web and API.
anton-l builds and maintains Wav2vec2 Large Xlsr 53 Mongolian, and it first shipped in 2022.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do