RinggAI/STT

Space

7

stars

5

commits

1

linked in READMEs

Apr 30, 2026

updated

asr
audio
bilingual
code-mixed
english
hindi
real-time
ringg
speech-to-text
static
transcription
Browse cluster: Automatic Speech Recognition (ASR)

README

Ringg Parrot STT V1

High-accuracy Hindi-English code-mixed speech-to-text for business voice applications.

Hugging Face Space Model Access

Overview

Ringg Parrot STT V1 is a proprietary automatic speech recognition system for Hindi, English, and Hindi-English code-mixed speech. It is designed for real-time voice products, business workflows, and production-grade speech-to-text use cases.

This Hugging Face Space is intended for product evaluation and release information. The model weights, training code, and internal implementation are not open sourced.

Access and Availability

  • Model weights: Not available for download from this repository.
  • Source code: Internal implementation is not open sourced.
  • Production and commercial access: Contact RinggAI at sales@ringg.ai.

Playground Usage

  1. Open the Ringg STT Playground.
  2. Upload an audio file or use the available streaming interface.
  3. Submit the audio for transcription.
  4. Review the generated Hindi, English, or code-mixed transcript.

SDK and Integration

  • Python SDK: ringglabs on PyPI
  • Pipecat: Highly compatible with Pipecat toolkit using built-in VAD events.
  • For integration details, refer to the SDK documentation.

Streaming Latency

Typical streaming latency is 60-80ms.

Benchmark Results

WER stands for Word Error Rate. Lower values indicate better transcription accuracy.

Original WER

DatasetRinggElevenLabsDeepgramSarvam
indictts11.5816.0613.6515.37
commonvoice14.3016.5920.0418.21
fleurs15.2011.9917.1416.00
kathbath11.7813.2415.9317.53
kathbath_noisy13.0913.1417.4416.19
mucs14.5511.6921.9716.72
Overall WER13.7913.0019.2316.72

Normalized WER

DatasetRinggElevenLabsDeepgramSarvam
indictts3.948.526.937.84
commonvoice6.3713.0214.8813.06
fleurs9.737.6711.359.54
kathbath7.1510.1511.3810.41
kathbath_noisy8.3710.0112.9811.78
mucs6.286.7512.077.58
Overall WER7.278.9412.369.76

Features

  • Hindi-English code-mixed speech recognition.
  • Real-time streaming transcription.
  • File-based transcription for common audio formats.
  • Low-latency inference for voice products.
  • Business-focused deployment and integration support.
  • Compatibility with modern voice-agent pipelines.

Supported Inputs

  • Languages: Hindi, English, and Hindi-English code-mixed speech.
  • Recommended audio: Clear speech with minimal background noise.
  • Sample rate: 16kHz or higher recommended.
  • Formats: WAV, MP3, FLAC, M4A, OGG, and OPUS.

Best Practices

  • Use clear audio captured close to the speaker.
  • Reduce background noise, echo, and overlapping speech where possible.
  • Use 16kHz or higher audio for best results.
  • Test representative production audio before deployment.

Use Cases

  • Voice assistants and AI agents.
  • Contact center transcription.
  • Meeting and conversation intelligence.
  • Voice search and commands.
  • Subtitling and content workflows.
  • Accessibility and documentation workflows.

Limitations

  • Accuracy may vary with noisy audio, overlapping speakers, strong accents, dialect variation, or low-quality recordings.
  • Performance can vary across domains, speaker profiles, and audio capture setups.
  • Very long files or unsupported encodings may require preprocessing.
  • The hosted demo is intended for evaluation and may not reflect every production deployment configuration.

Privacy and Data Notice

Audio handling may depend on the selected deployment, integration, and commercial terms. Review RinggAI privacy terms and deployment documentation before using the service with sensitive, regulated, or personally identifiable data.

Benchmark Dataset

RinggAI has released the ASR Benchmarking Open-Source Dataset, which includes benchmark audio/data and transcriptions generated by Ringg, ElevenLabs, Deepgram, and Sarvam.

Team

Built by the RinggAI Team.

Contributors

AnjulRSharma

3 commits

anjul1008

2 commits

RinggAI/STT

Space

7

stars

5

commits

1

linked in READMEs

Apr 30, 2026

updated

asr
audio
bilingual
code-mixed
english
hindi
real-time
ringg
speech-to-text
static
transcription
Browse cluster: Automatic Speech Recognition (ASR)

README

Ringg Parrot STT V1

High-accuracy Hindi-English code-mixed speech-to-text for business voice applications.

Hugging Face Space Model Access

Overview

Ringg Parrot STT V1 is a proprietary automatic speech recognition system for Hindi, English, and Hindi-English code-mixed speech. It is designed for real-time voice products, business workflows, and production-grade speech-to-text use cases.

This Hugging Face Space is intended for product evaluation and release information. The model weights, training code, and internal implementation are not open sourced.

Access and Availability

  • Model weights: Not available for download from this repository.
  • Source code: Internal implementation is not open sourced.
  • Production and commercial access: Contact RinggAI at sales@ringg.ai.

Playground Usage

  1. Open the Ringg STT Playground.
  2. Upload an audio file or use the available streaming interface.
  3. Submit the audio for transcription.
  4. Review the generated Hindi, English, or code-mixed transcript.

SDK and Integration

  • Python SDK: ringglabs on PyPI
  • Pipecat: Highly compatible with Pipecat toolkit using built-in VAD events.
  • For integration details, refer to the SDK documentation.

Streaming Latency

Typical streaming latency is 60-80ms.

Benchmark Results

WER stands for Word Error Rate. Lower values indicate better transcription accuracy.

Original WER

DatasetRinggElevenLabsDeepgramSarvam
indictts11.5816.0613.6515.37
commonvoice14.3016.5920.0418.21
fleurs15.2011.9917.1416.00
kathbath11.7813.2415.9317.53
kathbath_noisy13.0913.1417.4416.19
mucs14.5511.6921.9716.72
Overall WER13.7913.0019.2316.72

Normalized WER

DatasetRinggElevenLabsDeepgramSarvam
indictts3.948.526.937.84
commonvoice6.3713.0214.8813.06
fleurs9.737.6711.359.54
kathbath7.1510.1511.3810.41
kathbath_noisy8.3710.0112.9811.78
mucs6.286.7512.077.58
Overall WER7.278.9412.369.76

Features

  • Hindi-English code-mixed speech recognition.
  • Real-time streaming transcription.
  • File-based transcription for common audio formats.
  • Low-latency inference for voice products.
  • Business-focused deployment and integration support.
  • Compatibility with modern voice-agent pipelines.

Supported Inputs

  • Languages: Hindi, English, and Hindi-English code-mixed speech.
  • Recommended audio: Clear speech with minimal background noise.
  • Sample rate: 16kHz or higher recommended.
  • Formats: WAV, MP3, FLAC, M4A, OGG, and OPUS.

Best Practices

  • Use clear audio captured close to the speaker.
  • Reduce background noise, echo, and overlapping speech where possible.
  • Use 16kHz or higher audio for best results.
  • Test representative production audio before deployment.

Use Cases

  • Voice assistants and AI agents.
  • Contact center transcription.
  • Meeting and conversation intelligence.
  • Voice search and commands.
  • Subtitling and content workflows.
  • Accessibility and documentation workflows.

Limitations

  • Accuracy may vary with noisy audio, overlapping speakers, strong accents, dialect variation, or low-quality recordings.
  • Performance can vary across domains, speaker profiles, and audio capture setups.
  • Very long files or unsupported encodings may require preprocessing.
  • The hosted demo is intended for evaluation and may not reflect every production deployment configuration.

Privacy and Data Notice

Audio handling may depend on the selected deployment, integration, and commercial terms. Review RinggAI privacy terms and deployment documentation before using the service with sensitive, regulated, or personally identifiable data.

Benchmark Dataset

RinggAI has released the ASR Benchmarking Open-Source Dataset, which includes benchmark audio/data and transcriptions generated by Ringg, ElevenLabs, Deepgram, and Sarvam.

Team

Built by the RinggAI Team.

Contributors

AnjulRSharma

3 commits

anjul1008

2 commits