Push-to-talk
A global hotkey you choose. Hold to record, release to stop — text is typed straight into any app: editors, chat, email, IDE.
Hold a key, speak, release — text lands at your cursor. Your API key, your cloud, zero subscriptions.
A tiny native app that stays out of your way.
A global hotkey you choose. Hold to record, release to stop — text is typed straight into any app: editors, chat, email, IDE.
Paste your own API key and pay providers at cost. No middleman, no subscription, no usage caps invented by us.
sk-…e9f2 → stored in Keychain / DPAPI
Built with Tauri + Rust: an under 15 MB binary, instant start, negligible RAM. No Electron, no tray-resident bloat.
Audio goes only to the provider you configured. No telemetry, no accounts, no analytics — the code is open, check it.
One mental model on both desktops, localized interface in English and Russian, DMG and installer for your platform.
Your voice, your cloud, zero subscriptions.
No windows to open, no files to manage.
Press and hold your chosen shortcut in any application. The menu-bar indicator lights up.
Dictate a sentence, a paragraph or a commit message. Punctuation by voice, too.
Transcribed text is inserted at your cursor. That's the whole workflow.
Configure one or all three. VoiceDo never proxies your traffic.
Whisper and any endpoint that speaks the OpenAI audio API — including self-hosted servers on your LAN.
POST /v1/audio/transcriptions
Alibaba Cloud Qwen speech models — strong Russian and multilingual recognition at low cost.
qwen3-asr · paraformer
Zero-config fallback using Google's public transcription endpoint. Great for trying VoiceDo in one minute.
speech-api · instant
Free, open source, under 15 MB. Grab the build for your platform and dictate your first sentence in under a minute.
Demo builds are not digitally signed yet — macOS: right-click → "Open" on first launch; Windows SmartScreen may warn "unknown publisher", use "More info → Run".
VoiceDo is open source and always will be. Donations fuel development.