Voice Activity Detection
pyannote.audio
pyannote
pyannote-audio-pipeline
audio
voice
speech
speaker
speaker-diarization
speaker-change-detection
overlapped-speech-detection
Instructions to use BoKnows/vad-endpoint with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- pyannote.audio
How to use BoKnows/vad-endpoint with pyannote.audio:
from pyannote.audio import Pipeline pipeline = Pipeline.from_pretrained("BoKnows/vad-endpoint") # inference on the whole file pipeline("file.wav") # inference on an excerpt from pyannote.core import Segment excerpt = Segment(start=2.0, end=5.0) from pyannote.audio import Audio waveform, sample_rate = Audio().crop("file.wav", excerpt) pipeline({"waveform": waveform, "sample_rate": sample_rate}) - Notebooks
- Google Colab
- Kaggle
Download create_handler.ipynb from BoKnows/vad-endpoint: direct link, hf CLI and curl.
- Browser
- Download file 7.06 kB
-
https://huggingface.co/BoKnows/vad-endpoint/resolve/main/create_handler.ipynb
- Command line
-
hf download hf://BoKnows/vad-endpoint/create_handler.ipynb
-
curl -L -o create_handler.ipynb https://huggingface.co/BoKnows/vad-endpoint/resolve/main/create_handler.ipynb
7.06 kB