This is a large wav2vec2 model fine-tuned specifically for robust English phoneme recognition. Hosted on Hugging Face, it uses the Transformers library and is suitable for research and development in speech processing. The model has been trained to handle varied acoustic conditions.
Wav2vec2 Large Robust L2 English Phoneme Recognition is a Voice, TTS & speech product. Accurately transcribing English speech audio into phonemes for linguistic analysis or speech applications. It is built as an open-source project for AI researchers and developers. Wav2vec2 Large Robust L2 English Phoneme Recognition is open source under the Open Source license. The product ships for the web and API.
It is developed by slplab, and the product first shipped in 2025. Key capabilities include Automatic Speech Recognition, Phoneme Recognition, and Robust Audio Processing.
Latest indexed changes and source events
slplab/wav2vec2-large-robust-L2-english-phoneme-recognition verified by the PulseGate indexer
Other apps tracked under the same category.