Speech Recognition and On-Device Audio Processing

10 repos

Libraries and SDKs for speech-to-text and speech recognition, with emphasis on neural network-based models and on-device processing. The cluster centers around Deepgram's multi-language SDK implementations (Node.js, JavaScript, Go, Python, .NET) alongside complementary tools for audio processing, machine learning inference, and speech model deployment. Developers here work with both cloud-based and edge-deployed speech recognition systems, with particular attention to real-time transcription and latency-sensitive audio applications.

Python · 3
TypeScript · 3
C# · 2
Go · 1
JavaScript · 1
whisper ·3,742
whisper-ai ·3,648
docker-compose ·3,648
openai-whisper ·3,648
transcription ·3,648
docker ·3,648
openai-whisper-translation ·3,648
openai-api ·3,648
faster-whisper ·3,648
speech-to-text ·1,937