14 repos
Python libraries and models for identifying and tracking speakers in audio, including speaker diarization (who spoke when), speaker verification (confirming speaker identity), voice activity detection, and speaker embedding extraction. The cluster centers on pyannote-audio as a foundational framework, with specialized model variants (diarizen-wavlm) and complementary tools like diart and wespeaker providing alternative implementations, optimization approaches (ONNX), and production-ready pipelines for real-world audio processing tasks.