18 repos
Macoron/whisper.unity
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
751
60 commits
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
376
231 commits
ggerganov/whisper.cpp
Port of OpenAI's Whisper model in C/C++
53,731
4,480 commits
ggml-org/whisper.cpp
4,421 commits
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
394
368 commits
m1el/nemotron-asr.cpp
Nemotron ASR rewrite to GGML
22
74 commits
ayutaz/vokra
Speech-first inference runtime in Rust — TTS / ASR / speech-to-speech / VC / speaker ID / VAD. An…
10
198 commits
CrispStrobe/CrispASR
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral,…
641
3,603 commits
sandrohanea/whisper.net
Whisper.net. Speech to text made simple using Whisper Models
942
271 commits
soniqo/speech-core
On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent…
87
200 commits
k2-fsa/sherpa-ncnn
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn…
1,783
237 commits
wenet-e2e/wenet
Production First and Production Ready End-to-End Speech Recognition Toolkit
5,239
1,618 commits
kaldi-asr/kaldi
kaldi-asr/kaldi is the official location of the Kaldi project.
15,487
8,893 commits
mozilla/DeepSpeech
DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in…
26,776
2,727 commits
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
1,932
519 commits
mybigday/whisper.rn
React Native binding of whisper.cpp.
802
388 commits
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD…
14,840
2,037 commits
k2-fsa/sherpa
Speech-to-text server framework with next-gen Kaldi
993
631 commits