7 repos
Deep learning approaches to speech enhancement, synthesis, and real-time audio processing. The cluster centers on neural codec models and end-to-end speech systems built with PyTorch, including tools for training speech enhancement networks, full-duplex audio interaction, and audio compression. Repositories span from low-level audio DSP to high-level conversational AI applications, with a focus on practical speech quality improvements and efficient model deployment.