39 repos
pyannote/voice-activity-detection
No description
241
0 commits
pyannote/overlapped-speech-detection
64
zhoujiaming777/DIFFA-2
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
2
6 commits
pyannote/speaker-diarization-3.1
3,704
pyannote/speaker-diarization
1,331
pyannote/speaker-diarization-community-1
1,826
pyannote/brouhaha
31
pyannote/segmentation
697
laion/voice-tagging-whisper
Voice-Tagging Whisper
1
19 commits
neuphonic/distill-neucodec
25
audeering/wav2vec2-large-robust-12-ft-emotion-msp-dim
Model for Dimensional Speech Emotion Recognition based on Wav2vec 2.0
173
24 commits
pyannote/segmentation-3.0
1,851
neuphonic/neucodec
134
pyannote/embedding
238
OpenMOSS-Team/MOSS-Audio-8B-Instruct
MOSS-Audio
49
8 commits
OpenMOSS-Team/MOSS-Audio-4B-Instruct
82
7 commits
OpenMOSS-Team/MOSS-Audio-4B-Thinking
37
OpenMOSS-Team/MOSS-Audio-8B-Thinking
83
JusperLee/Hive
12 commits
neuphonic/neutts-air
894
neuphonic/neutts-air-q4-gguf
70
neuphonic/neutts-air-q8-gguf
43
3loi/SER-Odyssey-Baseline-WavLM-Arousal
The model was trained on [MSP-Podcast](https://ecs.utdallas.edu/research/researchlabs/msp-lab/MSP-P…
3loi/SER-Odyssey-Baseline-WavLM-Valence
11 commits
3loi/SER-Odyssey-Baseline-WavLM-Categorical
11
3loi/SER-Odyssey-Baseline-WavLM-Dominance
monishmal0204/nova-vad
NOVA-VAD
0
3 commits
3loi/SER-Odyssey-Baseline-WavLM-Multi-Attributes
13
33 commits
laion/vocalburst-locator
Vocal Burst Locator
5
18 commits
pyannote/wespeaker-voxceleb-resnet34-LM
Using this open-source model in production?
bodhan-ai/indic-transcribe-core
28
nvidia/audio-flamingo-next-hf
Model Overview
nvidia/audio-flamingo-next-captioner-hf
19
17 commits
nvidia/audio-flamingo-next-think-hf
9
16 commits
bodhan-ai/indic-transcribe-flex
21
laion/vocalburst-captioning-whisper
vocalburst-captioning-whisper
5 commits
laion/sound-effect-captioning-whisper
LAION Sound-Effect Captioning Whisper
mkrausio/EmoWhisper-AnS-Small-v0.1
Model Card for Model ID
4 commits
WhissleAI/stt_en_conformer_ctc_large_slurp