An open speech-recognition model fine-tuned from XLSR Wav2Vec2 for recognizing Japanese Hiragana. Developers can download and run it locally with the Transformers library for audio transcription workflows.
Wav2vec2 Large Xlsr Japanese Hiragana is a Speech to text project. It focuses on recognizing Japanese Hiragana from spoken audio without training a speech-recognition model from scratch. It is built as an open-source project for machine learning developers. The project is open source (Open Source). It ships for the command line, and it can be self-hosted.
It is developed by vumichien. It operates in a well-populated space: PulseGate tracks 5 similar projects. Among its 5 catalogued features are Japanese Hiragana Recognition, Automatic Speech Recognition, and Wav2Vec2 Architecture.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do