pyannote/voice-activity-detection is an MIT-licensed open-source speech activity detection model distributed through Hugging Face. Developers can load it with pyannote.audio and run inference locally on complete audio files or selected excerpts.
Voice Activity Detection sits in PulseGate's Speech to text category. It focuses on detecting when speech is present in audio recordings for downstream transcription and speaker analysis. It is built as an open-source project for speech and audio developers. Voice Activity Detection is open source under the MIT license. Voice Activity Detection is available on the web, API, and the command line, and it can be self-hosted.
pyannote builds and maintains Voice Activity Detection, and it first shipped in 2016. Development happens publicly on GitHub with 10.5k stars and 12 commits in the last 90 days. Among its 8 catalogued features are speech detection, audio segmentation, and file inference.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do