pyannote-speaker-diarization-3.1 is a pipeline model based on pyannote.audio for performing speaker diarization, voice activity detection, and overlapped speech detection. It can process entire audio files or specific excerpts and is designed for research and production audio analysis tasks. The model is hosted on Hugging Face and integrates directly with the pyannote.audio library.
Pyannote Speaker Diarization is a Voice, TTS & speech product. It focuses on determining who spoke when in multi-speaker audio recordings. It is built as an open-source project for developers. Pyannote Speaker Diarization is open source under the Open Source license. Pyannote Speaker Diarization is available on the web and API.
ivrit-ai builds and maintains Pyannote Speaker Diarization, and the product first shipped in 2023. Key capabilities include Speaker Diarization, Voice Activity Detection, and Overlapped Speech Detection.
Latest indexed changes and source events
ivrit-ai/pyannote-speaker-diarization-3.1 verified by the PulseGate indexer