5 repos
RicherMans/Dasheng
Source for the Interspeech 2024 Paper "Scaling up masked audio encoder learning for general audio…
87
20 commits
XiaoMi/dasheng
Official PyTorch code for Deep Audio-Signal Holistic Embeddings
206
jimbozhang/xares
A benchmark for evaluating audio encoders on various audio tasks.
56
86 commits
xiaomi-research/xares-llm
XARES-LLM
55
84 commits
vectominist/usad
Official implementation of "USAD: Universal Speech and Audio Representation via Distillation"
11
10 commits