Local speech-to-text powered by Whisper. Fast, private, and works offline.
C#
0
82 commits
updated Oct 3, 2026
Local speech-to-text for Windows, powered by Whisper. Press a global hotkey, speak, and your words land on the clipboard (and optionally type themselves at the cursor). Transcription runs entirely on your device by default. No account, no internet, no telemetry. Free and open source.
Talkty 1.5.0 is the latest stable release. See the release record for changes and verification.

On a Mac? There is a native Apple Silicon build too: talkty-mac (Metal accelerated, menu-bar app). Same idea, built the right way for each platform.
Most dictation tools send your microphone to someone else's server. That is a hard no for a lot of what people actually say out loud: client work, half-formed ideas, anything private. Whisper runs locally on your machine, new local recordings are held in memory, and no transcription data leaves the device unless you deliberately turn on a cloud feature.
It started as a tool for people who code by talking to an AI agent. Dictate a rambling thought, get clean text, paste it into Claude Code, Cursor, or a terminal. The optional Prompting mode goes one step further and rewrites that dictation into a structured prompt. But none of that is required. At its core it is a fast, private "hold a key, speak, get text" tool that works in any app.
Alt+Q) starts and stops
recording from any app. Text goes straight to the clipboard.kubectl, "post gres" becomes PostgreSQL). Fully editable in Settings.[MUSIC] and "Thanks for watching", and normalizes punctuation.
TalktySetup-*.exe from the
Releases page.Alt+Q, speak, press it again to stop. Your text is on the clipboard.Upgrading is the same: run the new installer over the old one. Your settings stay put.
Models download on demand from HuggingFace into %AppData%\Talkty\Models\.
| Tier | Model | Size | Best for |
|---|---|---|---|
| Fast | Tiny | 75 MB | Quick notes, simple phrases |
| Balanced | Small | 466 MB | Everyday English dictation |
| Balanced | Large v3 Turbo | 1.6 GB | 99+ languages, the all-round pick (recommended) |
| Accurate | Large v3 | 3.1 GB | Maximum accuracy |
Quantized "Lite" variants are available for CPU-only machines (smaller and lighter with a small accuracy trade). Pick any of them in Settings -> Local models.
Two features trade a little privacy for accuracy or convenience. Both are off by default and both run through a single OpenRouter API key that is stored encrypted on your device (Windows DPAPI), never in plain text.

Leave both off and Talkty stays 100% local.
For the shortest plain-dictation delay, leave Prompting off. Optional prompt planning and fidelity checks are documented in planning and fidelity; they can add paid text requests when configured. Planning is an advanced setting and defaults off.
Start transcription during pauses is enabled by default under Settings > Behavior. With a cloud model selected, audio can be sent before you press Stop. At most one extra early pass may be billed, even if you cancel or continue speaking. Disable this setting to send cloud audio only after stopping.
Local transcription is fully private:
If you turn on Cloud transcription or Prompting, the relevant audio or text is sent
to OpenRouter and its model providers. A cloud backup can send the same recording
to another provider when the primary fails. Your API key is encrypted on disk.
Cloud recordings are saved under %AppData%\Talkty\Recovery, encrypted with Windows
DPAPI for your Windows account, at Stop before the final upload. An early request
during a pause can precede that save. The recovery copy is removed after
the transcript is saved to history or when you choose Discard. Failed takes survive
restarts; switching to a local model does not delete them. Turn cloud features off
to keep new transcription requests offline.
History and diagnostic logs stay on your PC; logs may contain dictated text. Model downloads and update checks use the network. Optional command mode passes transcript and window information to the configured local service, which controls any further actions or network access.
Open Settings from the gear icon or by right-clicking the tray icon.
| Setting | What it does |
|---|---|
| Model | Which Whisper model to use, local or cloud |
| Microphone | Audio input device, with a built-in test |
| Hotkey | Global shortcut (default Alt+Q) |
| Language | Force a language or auto-detect |
| GPU | Use CUDA or Vulkan acceleration when available |
| Auto-paste | Insert the text at the cursor after transcription |
| Volume ducking | Lower other audio while recording |
| Vocabulary | Custom coding terms and text replacements |
| API key | OpenRouter key for Cloud and Prompting (encrypted) |
| Cloud backup | Automatic fallback model for temporary cloud failures, or Off |
| Transcription during pauses | Prepare a whole-take result early; cloud can bill one extra pass |
| Prompting | Rewrite ordinary dictation into an agent prompt; adds a paid request and delay |
| Command mode | Separate shortcut for a configured local command service; off by default |
Requirements: .NET 8 SDK, Windows 10 or 11, and (for the installer) Inno Setup 6.
# Build and run
dotnet build Talkty.App
dotnet run --project Talkty.App
# Checked release build and installer, in a fresh publish directory
pwsh -File installer/build.ps1
Run the tests with dotnet test. The version number lives in a single place,
version.txt, and flows into the assembly and the installer automatically.
The script keeps native runtime files alongside the executable, verifies the
payload, and writes SHA-256 hashes. See the release guide.
Talkty.App/
Models/ App settings, model profiles, vocabulary defaults
Services/ Audio capture, transcription, hotkey, auto-paste, clipboard
Engines/ Whisper (CUDA/Vulkan/CPU), SherpaOnnx, OpenRouter cloud
ViewModels/ Recording state machine and settings logic (MVVM)
Views/ Main window, settings, overlay pill, onboarding
Talkty.Tests/ xUnit tests for recording, output, cloud recovery, prompts and UI
installer/ Inno Setup script
docs/ PROMPTING.md and assets
A deeper tour of the architecture lives in docs/ARCHITECTURE.md.
.NET 8 and WPF, MVVM via CommunityToolkit.Mvvm, Whisper.net for transcription, NAudio for capture, and Hardcodet.NotifyIcon.Wpf for the tray.
Issues and pull requests are welcome. Start with CONTRIBUTING.md for how to build, what gets tested, and the conventions to match. Security policy and how to report a vulnerability: SECURITY.md.
C#
87.8%
PowerShell
5.6%
TypeScript
2.0%
HTML
2.0%
Local speech-to-text powered by Whisper. Fast, private, and works offline.
C#
0
82 commits
updated Oct 3, 2026
Local speech-to-text for Windows, powered by Whisper. Press a global hotkey, speak, and your words land on the clipboard (and optionally type themselves at the cursor). Transcription runs entirely on your device by default. No account, no internet, no telemetry. Free and open source.
Talkty 1.5.0 is the latest stable release. See the release record for changes and verification.

On a Mac? There is a native Apple Silicon build too: talkty-mac (Metal accelerated, menu-bar app). Same idea, built the right way for each platform.
Most dictation tools send your microphone to someone else's server. That is a hard no for a lot of what people actually say out loud: client work, half-formed ideas, anything private. Whisper runs locally on your machine, new local recordings are held in memory, and no transcription data leaves the device unless you deliberately turn on a cloud feature.
It started as a tool for people who code by talking to an AI agent. Dictate a rambling thought, get clean text, paste it into Claude Code, Cursor, or a terminal. The optional Prompting mode goes one step further and rewrites that dictation into a structured prompt. But none of that is required. At its core it is a fast, private "hold a key, speak, get text" tool that works in any app.
Alt+Q) starts and stops
recording from any app. Text goes straight to the clipboard.kubectl, "post gres" becomes PostgreSQL). Fully editable in Settings.[MUSIC] and "Thanks for watching", and normalizes punctuation.
TalktySetup-*.exe from the
Releases page.Alt+Q, speak, press it again to stop. Your text is on the clipboard.Upgrading is the same: run the new installer over the old one. Your settings stay put.
Models download on demand from HuggingFace into %AppData%\Talkty\Models\.
| Tier | Model | Size | Best for |
|---|---|---|---|
| Fast | Tiny | 75 MB | Quick notes, simple phrases |
| Balanced | Small | 466 MB | Everyday English dictation |
| Balanced | Large v3 Turbo | 1.6 GB | 99+ languages, the all-round pick (recommended) |
| Accurate | Large v3 | 3.1 GB | Maximum accuracy |
Quantized "Lite" variants are available for CPU-only machines (smaller and lighter with a small accuracy trade). Pick any of them in Settings -> Local models.
Two features trade a little privacy for accuracy or convenience. Both are off by default and both run through a single OpenRouter API key that is stored encrypted on your device (Windows DPAPI), never in plain text.

Leave both off and Talkty stays 100% local.
For the shortest plain-dictation delay, leave Prompting off. Optional prompt planning and fidelity checks are documented in planning and fidelity; they can add paid text requests when configured. Planning is an advanced setting and defaults off.
Start transcription during pauses is enabled by default under Settings > Behavior. With a cloud model selected, audio can be sent before you press Stop. At most one extra early pass may be billed, even if you cancel or continue speaking. Disable this setting to send cloud audio only after stopping.
Local transcription is fully private:
If you turn on Cloud transcription or Prompting, the relevant audio or text is sent
to OpenRouter and its model providers. A cloud backup can send the same recording
to another provider when the primary fails. Your API key is encrypted on disk.
Cloud recordings are saved under %AppData%\Talkty\Recovery, encrypted with Windows
DPAPI for your Windows account, at Stop before the final upload. An early request
during a pause can precede that save. The recovery copy is removed after
the transcript is saved to history or when you choose Discard. Failed takes survive
restarts; switching to a local model does not delete them. Turn cloud features off
to keep new transcription requests offline.
History and diagnostic logs stay on your PC; logs may contain dictated text. Model downloads and update checks use the network. Optional command mode passes transcript and window information to the configured local service, which controls any further actions or network access.
Open Settings from the gear icon or by right-clicking the tray icon.
| Setting | What it does |
|---|---|
| Model | Which Whisper model to use, local or cloud |
| Microphone | Audio input device, with a built-in test |
| Hotkey | Global shortcut (default Alt+Q) |
| Language | Force a language or auto-detect |
| GPU | Use CUDA or Vulkan acceleration when available |
| Auto-paste | Insert the text at the cursor after transcription |
| Volume ducking | Lower other audio while recording |
| Vocabulary | Custom coding terms and text replacements |
| API key | OpenRouter key for Cloud and Prompting (encrypted) |
| Cloud backup | Automatic fallback model for temporary cloud failures, or Off |
| Transcription during pauses | Prepare a whole-take result early; cloud can bill one extra pass |
| Prompting | Rewrite ordinary dictation into an agent prompt; adds a paid request and delay |
| Command mode | Separate shortcut for a configured local command service; off by default |
Requirements: .NET 8 SDK, Windows 10 or 11, and (for the installer) Inno Setup 6.
# Build and run
dotnet build Talkty.App
dotnet run --project Talkty.App
# Checked release build and installer, in a fresh publish directory
pwsh -File installer/build.ps1
Run the tests with dotnet test. The version number lives in a single place,
version.txt, and flows into the assembly and the installer automatically.
The script keeps native runtime files alongside the executable, verifies the
payload, and writes SHA-256 hashes. See the release guide.
Talkty.App/
Models/ App settings, model profiles, vocabulary defaults
Services/ Audio capture, transcription, hotkey, auto-paste, clipboard
Engines/ Whisper (CUDA/Vulkan/CPU), SherpaOnnx, OpenRouter cloud
ViewModels/ Recording state machine and settings logic (MVVM)
Views/ Main window, settings, overlay pill, onboarding
Talkty.Tests/ xUnit tests for recording, output, cloud recovery, prompts and UI
installer/ Inno Setup script
docs/ PROMPTING.md and assets
A deeper tour of the architecture lives in docs/ARCHITECTURE.md.
.NET 8 and WPF, MVVM via CommunityToolkit.Mvvm, Whisper.net for transcription, NAudio for capture, and Hardcodet.NotifyIcon.Wpf for the tray.
Issues and pull requests are welcome. Start with CONTRIBUTING.md for how to build, what gets tested, and the conventions to match. Security policy and how to report a vulnerability: SECURITY.md.
C#
87.8%
PowerShell
5.6%
TypeScript
2.0%
HTML
2.0%