10 repos
Libraries and models for detecting speech presence in audio streams, with emphasis on real-time voice activity detection (VAD) across multiple speaker scenarios. The cluster centers on pre-trained VAD models (particularly Silero VAD variants and SortFormer-based diarization models) and toolkits that enable speech activity detection for voice command recognition, speaker diarization, and audio preprocessing pipelines.