This fine-tuned wav2vec2 model classifies audio for speaker age groups and gender. It is designed for audio analysis tasks and can be used with the Hugging Face Transformers library. The model was developed by audEERING and released under a CC-BY-NC-SA-4.0 license for research and non-commercial applications.
In the Voice, TTS & speech space, Wav2vec2 Large Robust 6 Ft Age Gender takes a focused approach. Accurately determining speaker age and gender directly from raw audio waveforms. Wav2vec2 Large Robust 6 Ft Age Gender is an open-source project aimed at developers. Wav2vec2 Large Robust 6 Ft Age Gender is open source under the MIT license. It runs on the web and API, and it can be self-hosted.
Behind Wav2vec2 Large Robust 6 Ft Age Gender is audEERING GmbH, and it first shipped in 2023. Development happens publicly on GitHub with 56 stars. Key capabilities include Audio Classification, Age Recognition, and Gender Recognition. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do