Open source ยท AGPL-3.0 ยท under 15 MB

Speak. It's written.

Native push-to-talk dictation for macOS and Windows. Hold a key, talk, release โ€” your words land at the cursor. Bring your own API key: your voice, your cloud, zero subscriptions.

Free forever. No account required. Works with OpenAI, Qwen and Google.

VoiceDo โ€” voice dictation on a laptop, dark theme
โŒ˜ Push-to-talk ๐Ÿ”‘ Bring your own key ๐ŸŽ macOS ๐ŸชŸ Windows ๐Ÿ“ฆ under 15 MB native binary ๐ŸŒ EN / RU interface

Everything dictation needs. Nothing it doesn't.

A tiny native app that stays out of your way.

โŒจ๏ธ

Push-to-talk

A global hotkey you choose. Hold to record, release to stop โ€” text is typed straight into any app: editors, chat, email, IDE.

๐Ÿ”‘

Bring your own key

Paste your own API key and pay providers at cost. No middleman, no subscription, no usage caps invented by us.

๐Ÿง 

Three provider families

OpenAI-compatible endpoints (incl. Whisper), Qwen / DashScope, and Google speech โ€” switchable per language and budget.

๐Ÿ“ฆ

Featherweight native

Built with Tauri + Rust: an under 15 MB binary, instant start, negligible RAM. No Electron, no tray-resident bloat.

๐Ÿ”’

Your audio, your rules

Audio goes only to the provider you configured. No telemetry, no accounts, no analytics โ€” the code is open, check it.

๐ŸŒ

macOS & Windows

One mental model on both desktops, localized interface in English and Russian, DMG and installer for your platform.

Three seconds from speech to text

No windows to open, no files to manage.

1

Hold the hotkey

Press and hold your chosen shortcut in any application. The menu-bar indicator lights up.

2

Speak naturally

Dictate a sentence, a paragraph or a commit message. Punctuation by voice, too.

3

Release โ€” done

Transcribed text is inserted at your cursor. That's the whole workflow.

Bring a key, pick a brain

Configure one or all three. VoiceDo never proxies your traffic.

Default

OpenAI-compatible

Whisper and any endpoint that speaks the OpenAI audio API โ€” including self-hosted servers on your LAN.

POST /v1/audio/transcriptions
Best for RU

Qwen ยท DashScope

Alibaba Cloud Qwen speech models โ€” strong Russian and multilingual recognition at low cost.

qwen3-asr ยท paraformer
No key needed

Google Speech

Zero-config fallback using Google's public transcription endpoint. Great for trying VoiceDo in one minute.

speech-api ยท instant

Type less. Say more.

Free, open source, under 15 MB. Grab the build for your platform and dictate your first sentence in under a minute.

Demo builds are not digitally signed yet โ€” macOS: right-click โ†’ "Open" on first launch; Windows SmartScreen may warn "unknown publisher", use "More info โ†’ Run".

v0.2.0 ยท 2026-09-04 ยท now on GitHub Releases ยท signed builds coming next

Like it? Keep it free.

VoiceDo is open source and always will be. Donations fuel development.