This model is trained to perform binary audio classification to detect whether a speaker is male or female using the Common Voice dataset. It is based on the wav2vec2 architecture and can be used with the Transformers pipeline for audio-classification. The model is suitable for voice analysis, demographic analysis, and accessibility applications.
Common Voice Gender Detection is a Voice, TTS & speech project. Automatically determining speaker gender (male/female) from audio samples. It is built as an open-source project for developers. Common Voice Gender Detection is open source under the Open Source license. It runs on the web, the command line, and API.
Behind Common Voice Gender Detection is prithivMLmods, and it first shipped in 2024. Key capabilities include Audio Classification, Gender Detection, and wav2Vec2.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do