pyannote.audio Neural Speaker Diarization Toolkit

pyannote.audio is an open-source Python toolkit for speaker diarization built on PyTorch. It provides state-of-the-art pretrained models and pipelines for speech activity detection, speaker segmentation, overlapped speech detection, and speaker embedding.

agentskillexchange Updated 28 repo stars

File contents

pyannote.audio Neural Speaker Diarization Toolkit

pyannote.audio is an open-source Python toolkit for speaker diarization built on PyTorch. It provides state-of-the-art pretrained models and pipelines for speech activity detection, speaker segmentation, overlapped speech detection, and speaker embedding.

Installation

Use the upstream install or setup path that matches your environment:

  • pip install -e .[dev,testing]

Requirements and caveats from upstream:

  • pyannote.audio is an open-source toolkit written in Python for speaker diarization. Based on PyTorch machine learning framework, it comes with state-of-the-art [pretrained models and pipelines](...
  • :snake: Python-first API
  • python

Basic usage or getting-started notes:

Source

agentskillexchange/skills/tree/main/skills/pyannote-audio-speaker-diarization-toolkit commit b99a053a54

Frequently asked questions

npx skillmds@latest add agentskillexchange/pyannote-audio-neural-speaker-diarization-toolkit