24 repos
timoncool/HiggsAudio-Studio
Portable Windows TTS — Higgs Audio v3 + AI text director, podcast & audiobook multi-voice.…
79
40 commits
timoncool/dub-studio
Free offline AI video dubbing studio for Windows — voice cloning, translation, subtitles &…
118
312 commits
0xShug0/audio.cpp
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD,…
2,515
564 commits
debpalash/VoiceStudio
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design,…
21,995
1,785 commits
Saganaki22/Higgs-Audio-v3-Studio
Windows / Linux desktop app built with Rust/Tauri for local Higgs Audio v3 TTS inference + local…
32
49 commits
EveningStudy/asmr-dubber
音视频字幕及配音工具:支持 ASR(语音识别)、台本导入、AI 翻译、双语字幕、音色克隆、TTS(语音合成)与混音以得到双语音频。
171
66 commits
adzetto/reading-library-tts
Colab uretim hatti: reading-library belgelerini Step-Audio-EditX ile seslendirir. Whisper ile…
0
11 commits
soniqo/speech-swift
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by…
1,165
651 commits
mazzasaverio/youtube-auto-dub
Local-first, open-source YouTube dubbing: clones the original voice into another language,…
70
107 commits
dudarenok-maker/Castwright
A local, multi-voice audiobook generator: an LLM casts every character in its own voice, kept…
15
9,541 commits
soniqo/speech-android
On-device speech SDK for Android — ASR, TTS, VAD, and noise cancellation powered by ONNX Runtime…
152
140 commits
ayutaz/vokra
Speech-first inference runtime in Rust — TTS / ASR / speech-to-speech / VC / speaker ID / VAD. An…
9
198 commits
DePasqualeOrg/mlx-audio-plus
Python tools for text to speech (TTS), speech to text (STT), and speech to speech (STS) powered by…
47
311 commits
sgl-project/sglang-omni
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified…
1,141
862 commits
shawnrushefsky/talky-talky
MCP server for Audio Generation and Analysis with a Variety of Open Models.
2
60 commits
AutoArk/GPA
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
3,122
38 commits
tencent/AuK
No description
91
10 commits
tencent/AuK-Flash
unslothai/unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4,…
75,967
7,907 commits
drbaph/Higgs-Audio-v3-Studio
7
34 commits
xirtus/MOSSlanding
The best voice-cloning TTS on the planet. Clone any voice, 31 languages, runs 100% offline. Free…
1
15 commits
FlorianEagox/WeeaBlind
A program to dub non-english media with modern AI speech synthesis, diarization, and voice cloning!
323
111 commits
espeak-ng/espeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and…
6,826
5,865 commits
lucadellalib/audiocodecs
A collections of audio codecs with a standardized API
44
22 commits