pyannote-speaker-diarization-3.1 is a pipeline model based on pyannote.audio for performing speaker diarization, voice activity detection, and overlapped speech detection. It can process entire audio files or specific excerpts and is designed for research and production audio analysis tasks. The model is hosted on Hugging Face and integrates directly with the pyannote.audio library.
In the Speech to text space, Pyannote Speaker Diarization takes a focused approach. It focuses on determining who spoke when in multi-speaker audio recordings. Pyannote Speaker Diarization is an open-source project aimed at developers. The project is open source (Open Source). Pyannote Speaker Diarization is available on the web and API.
ivrit-ai builds and maintains Pyannote Speaker Diarization. Among its 3 catalogued features are Speaker Diarization, Voice Activity Detection, and Overlapped Speech Detection.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do