This model is a fine-tuned or adapted version of the pyannote.audio segmentation model (version 3.0). It is designed for tasks including speaker diarization, voice activity detection, overlapped speech detection, and speaker change detection. The model is hosted on Hugging Face under the ivrit-ai organization and is available for PyTorch inference.
In the Speech to text space, Pyannote Segmentation takes a focused approach. Accurately detecting and segmenting individual speakers in audio recordings. Pyannote Segmentation is an open-source project aimed at speech technology developers. Pyannote Segmentation is open source under the Open Source license. It runs on the web and API.
It is developed by ivrit.ai. Key capabilities include Speaker Diarization, Voice Activity Detection, and Speaker Change Detection.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do