This open-source wav2vec2 model performs automatic speech recognition for Portuguese audio. Developers can load it locally with Transformers and PyTorch or use compatible inference tooling for transcription workflows.
Wav2vec2 Large Xlsr 53 Portuguese sits in PulseGate's Speech to text category. It focuses on transcribing Portuguese speech into text without building an acoustic model from scratch. Wav2vec2 Large Xlsr 53 Portuguese is an open-source project aimed at machine learning developers and speech researchers. Wav2vec2 Large Xlsr 53 Portuguese is open source under the MIT license. It ships for the web and API, and it can be self-hosted.
Jonatas Grosman builds and maintains Wav2vec2 Large Xlsr 53 Portuguese, and it first shipped in 2021. The project is developed in the open on GitHub with 206 stars. It operates in a well-populated space: PulseGate tracks 7 similar projects. Among its 8 catalogued features are automatic speech recognition, portuguese transcription, and transformers integration.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do