1 个仓库
Metrics and tools for assessing the accuracy of boundaries in audio segmentation.
Distinct from Audio Segmentation Utilities: Focuses on evaluation of boundaries rather than the utility of splitting the audio
Explore 1 awesome GitHub repository matching graphics & multimedia · Segmentation Evaluation. Refine with filters or upvote what's useful.
Pyannote.audio is a PyTorch toolkit for speaker diarization, speaker identification, and speech activity detection. Its primary purpose is to partition audio recordings into segments and assign each segment to a specific speaker identity to determine who spoke when. The project includes a framework for classifying speaker identities and a pipeline for distinguishing human speech from background noise. It provides specialized tools for handling symmetric-overlap speech, where multiple speakers talk simultaneously, and employs learnable band-pass filters for raw waveform feature extraction. Th
Assesses the accuracy of detected speaker boundaries to determine how precisely speech turns are divided.