wavekat/wavekat-lab

Developer experimentation tools for the WaveKat libraries. Includes vad-lab, a web-based tool for testing and comparing VAD backends side by side.

11

stars

51

commits

Jupyter Notebook

primary language

Jun 4, 2026

updated

audio
audio-processing
developer-tools
rust
speech-detection
vad
voice
voice-ai
wavekat
Browse cluster: Voice Activity Detection (VAD)

README

WaveKat Lab

CI Release Please DeepWiki

A research repo for the WaveKat project — interactive tools and Jupyter notebooks for working with audio models (VAD, turn detection, voice datasets, and more).

[!WARNING] Early development. Things may change.

What's In Here

wavekat-lab/
├── tools/
│   ├── audio-lab/     Real-time VAD + Turn Detection + ASR comparison app (Rust + React)
│   └── cv-explorer/   Mozilla Common Voice dataset browser (Cloudflare Workers + React)
├── notebooks/         Jupyter notebooks (training, validation, dataset splits)
└── docs/              Plans and design docs

Each tool is self-contained — its own Makefile, lockfiles, and build setup live inside its folder.

Tools

Audio Labtools/audio-lab/

Web app for testing and comparing WaveKat library backends side by side in real time. Live mic capture, WAV upload, multi-config fan-out, VAD-gated pipeline mode, live ASR transcripts, waveform + spectrogram + probability timelines.

Backends: webrtc-vad, silero-vad, ten-vad, firered-vad, pipecat smart-turn, sherpa-onnx ASR. Details →

Common Voice Explorertools/cv-explorer/

Web app for browsing and playing audio clips from the Mozilla Common Voice dataset. Filter by locale, split, demographics, and search sentences — with waveform playback powered by WaveSurfer.js. Built on Cloudflare Workers + D1 + R2. Details →

Live: https://commonvoice-explorer.wavekat.com/

Notebooks

notebooks/ is the home for Jupyter notebooks covering training, validation, and dataset-splitting workflows. Python env is managed by uv.

make setup-notebooks   # one-time: uv sync the notebook env
make lab               # start Jupyter Lab on notebooks/

Repo Layout Conventions

  • Per-tool Makefilestools/<name>/Makefile owns dev/build/CI for that tool. Run cd tools/<name> && make help to see what's there.
  • Root Makefile — repo-wide only: setup, lab, ci, and per-tool CI delegators.
  • No shared Cargo workspace at root — each Rust tool keeps its own Cargo.toml / Cargo.lock / target/ inside its folder.

Videos

VideoDescription
Common Voice Explorer DemoExploring Mozilla Common Voice with Common Voice Explorer
Introducing Common Voice Explorer — browse and listen to 1.8M+ real voice clips from the Mozilla Common Voice dataset.
Pipecat Smart Turn Visual TestTesting Pipecat Smart Turn with WaveKat Lab
Visual test of Pipecat Smart Turn v3 — live recording and VAD-gated pipeline mode simulating production workflows.
FireRed VAD ShowdownAdding FireRedVAD as the 4th backend
Benchmarking Xiaohongshu's FireRedVAD against Silero, TEN VAD, and WebRTC across accuracy and latency.
VAD Lab DemoVAD Lab: Real-time multi-backend comparison
Live demo of VAD Lab comparing WebRTC, Silero, and TEN VAD side by side with real-time waveform visualization.

About WaveKat

wavekat-lab is part of WaveKat, an open-source ecosystem for building real-time voice pipelines. This repo holds interactive tools and notebooks for experimenting with the audio models — VAD, turn detection, and more — that power those pipelines.

See wavekat.com for the full project.

Stars

wavekat/wavekat-lab stars

License

Licensed under Apache 2.0.

Copyright 2026 WaveKat.

Contributors

wavekat/wavekat-lab

Developer experimentation tools for the WaveKat libraries. Includes vad-lab, a web-based tool for testing and comparing VAD backends side by side.

11

stars

51

commits

Jupyter Notebook

primary language

Jun 4, 2026

updated

audio
audio-processing
developer-tools
rust
speech-detection
vad
voice
voice-ai
wavekat
Browse cluster: Voice Activity Detection (VAD)

README

WaveKat Lab

CI Release Please DeepWiki

A research repo for the WaveKat project — interactive tools and Jupyter notebooks for working with audio models (VAD, turn detection, voice datasets, and more).

[!WARNING] Early development. Things may change.

What's In Here

wavekat-lab/
├── tools/
│   ├── audio-lab/     Real-time VAD + Turn Detection + ASR comparison app (Rust + React)
│   └── cv-explorer/   Mozilla Common Voice dataset browser (Cloudflare Workers + React)
├── notebooks/         Jupyter notebooks (training, validation, dataset splits)
└── docs/              Plans and design docs

Each tool is self-contained — its own Makefile, lockfiles, and build setup live inside its folder.

Tools

Audio Labtools/audio-lab/

Web app for testing and comparing WaveKat library backends side by side in real time. Live mic capture, WAV upload, multi-config fan-out, VAD-gated pipeline mode, live ASR transcripts, waveform + spectrogram + probability timelines.

Backends: webrtc-vad, silero-vad, ten-vad, firered-vad, pipecat smart-turn, sherpa-onnx ASR. Details →

Common Voice Explorertools/cv-explorer/

Web app for browsing and playing audio clips from the Mozilla Common Voice dataset. Filter by locale, split, demographics, and search sentences — with waveform playback powered by WaveSurfer.js. Built on Cloudflare Workers + D1 + R2. Details →

Live: https://commonvoice-explorer.wavekat.com/

Notebooks

notebooks/ is the home for Jupyter notebooks covering training, validation, and dataset-splitting workflows. Python env is managed by uv.

make setup-notebooks   # one-time: uv sync the notebook env
make lab               # start Jupyter Lab on notebooks/

Repo Layout Conventions

  • Per-tool Makefilestools/<name>/Makefile owns dev/build/CI for that tool. Run cd tools/<name> && make help to see what's there.
  • Root Makefile — repo-wide only: setup, lab, ci, and per-tool CI delegators.
  • No shared Cargo workspace at root — each Rust tool keeps its own Cargo.toml / Cargo.lock / target/ inside its folder.

Videos

VideoDescription
Common Voice Explorer DemoExploring Mozilla Common Voice with Common Voice Explorer
Introducing Common Voice Explorer — browse and listen to 1.8M+ real voice clips from the Mozilla Common Voice dataset.
Pipecat Smart Turn Visual TestTesting Pipecat Smart Turn with WaveKat Lab
Visual test of Pipecat Smart Turn v3 — live recording and VAD-gated pipeline mode simulating production workflows.
FireRed VAD ShowdownAdding FireRedVAD as the 4th backend
Benchmarking Xiaohongshu's FireRedVAD against Silero, TEN VAD, and WebRTC across accuracy and latency.
VAD Lab DemoVAD Lab: Real-time multi-backend comparison
Live demo of VAD Lab comparing WebRTC, Silero, and TEN VAD side by side with real-time waveform visualization.

About WaveKat

wavekat-lab is part of WaveKat, an open-source ecosystem for building real-time voice pipelines. This repo holds interactive tools and notebooks for experimenting with the audio models — VAD, turn detection, and more — that power those pipelines.

See wavekat.com for the full project.

Stars

wavekat/wavekat-lab stars

License

Licensed under Apache 2.0.

Copyright 2026 WaveKat.

Contributors

Languages

Jupyter Notebook

97.4%

TypeScript

1.4%