This is a fine-tuned version of the large XLS-R wav2vec2 model adapted specifically for Catalan speech recognition. Available on Hugging Face, it integrates with the Transformers library. The model was developed to support the Catalan language community and researchers working on low-resource language technologies.
In the Speech to text space, Wav2vec2 Large Xlsr Catala takes a focused approach. It focuses on providing accurate speech-to-text capabilities for the Catalan language using open models. Wav2vec2 Large Xlsr Catala is an open-source project aimed at AI developers and researchers. Wav2vec2 Large Xlsr Catala is open source under the Open Source license. It ships for the web and API.
It is developed by Softcatalà, and it first shipped in 2021. The project is developed in the open on GitHub with 12 stars. Among its 3 catalogued features are Automatic Speech Recognition, Catalan Language, and Transformers Support.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do