Console Download free
Local-first · Privacy by design

Turn every spoken word into clean, usable text

Voice rewrites productivity.声音,重写生产力。

TakeVox is a local-first speech-to-text and meeting assistant: an offline dictation engine, a live floating transcript, AI polishing and translation, and speaker-separated meeting minutes — with your audio staying on your device by default.

Free · no sign-up · data stays on your device

Latest release:

Free download · no sign-up Local-first · audio stays on device ZH / EN / JA / KO / YUE recognition macOS / Windows / Linux

Why TakeVox?

Four everyday pain points — one local-first answer.

Typing is too slow to keep up
Speaking is 3–4× faster than typing. Press the hotkey, speak, release — text lands, your flow never breaks.
Meeting notes always miss something
Speaker separation + auto minutes + action items. Never replay the recording again.
Uploading audio feels wrong
Audio stays on your device; local mode is fully offline; cloud goes through your own key — never through TakeVox servers.
Transcription bills by the minute
Local mode is free and unlimited; bring your own API key and control the cost.

How it works

From hotkey press to inserted text in under three seconds

1

Press & speak

In any app, press the global hotkey to start recording.

2

Live transcript

Watch the recognition appear in the overlay as you speak.

3

Stop to insert

Press the hotkey again — the final text lands at your cursor. ESC cancels.

Features

From dictation to meetings — the full path from speech to productivity.

Stop typing. Start thriving.不再打字,即刻迈向高效。

Meeting transcripts & minutes

Record or import audio for speaker separation, auto-generated minutes and action items with Markdown export.

转写分离 对齐整理 纪要行动项

Clipboard history

Automatically captures copied content with favorites, full-text search and instant paste.

TakeVox 会议纪要 v0.2⌘⇧V
离线 听写 使用说明
剪贴板 全文搜索

Offline live dictation

A built-in local speech engine turns voice into text in real time, no network required.

Multilingual recognition & translation

Auto-detects Chinese, English and more; transcripts can be translated in full.

Global hotkeys

Start/stop dictation in any app with one shortcut; text is inserted at the cursor.

Floating live transcript

A translucent overlay shows transcription as you speak, with cool/warm themes.

AI polish & translation

Polish wording, rewrite style, or translate into many languages with a local or cloud LLM.

Professional glossary

Import industry term lists (CSV/JSON): misreadings are auto-corrected, domain terms are recognized right on the first pass, and cloud transcription syncs the same hotwords.

别特 → 比特纠错对
RSC → React Server Components词表
CSV / JSON 批量导入…

Local-first privacy

Audio and transcripts stay on your device; cloud transcription uses your own subscription key directly.

Who uses TakeVox?

Writing, meetings, study, development — anyone who turns speech into text.

Writing & interviews

Journalists, authors, creators: interviews become drafts directly, ideas captured instantly.

Meetings & teamwork

Real-time transcription, auto minutes and action items at a glance.

Study & lectures

Students and lifelong learners: transcribe lectures live, translate on the fly.

Development & notes

Developers: dictate design docs; clipboard history recalls code snippets.

Three modes, pick what fits

From “most private” to “zero setup”, TakeVox is the same app — only how your speech is processed differs.

Speak. Instant Produced.开口即成文。

Local mode

Privacy-first
  • Fully offline — audio and transcripts never leave your device
  • Works without an account or network
  • Runs entirely on-device after the first model download
For sensitive content and air-gapped environments

Own API mode

Accuracy & control
  • Bring your own LLM / transcription service
  • OpenAI-compatible endpoint, free choice of model & parameters
  • Share one service setup across your team
For professional workloads that demand higher accuracy

Subscription mode

Zero setup
Pro
  • No model management — cloud transcription instantly
  • Extended audio and speaker separation
  • Synced across devices, pick up where you left off
For convenience and always-latest capabilities

Local-first vs. cloud transcription

Same speech-to-text — different data ownership, cost and availability.

Dimension TakeVox local mode Traditional cloud STT
Data ownership Audio & transcripts stay on device Uploaded to provider servers
Offline use Fully offline, no network needed Requires a connection
Pricing Free locally, unlimited Per-minute / per-usage
Privacy Stays on device by default Depends on provider policy
Languages ZH / EN / JA / KO / YUE Varies by provider
Security Built-in protection, encrypted by default Varies by provider

Enterprise plan

Deploy speech-to-text inside your organization: data stays internal, terminology matches your business, and privacy stays under your control.

On-premises deployment

Private · integrable
  • Deploy on intranet or private cloud — data never leaves your boundary
  • Integrate with existing OA, meeting and document systems
  • Unified accounts and role-based access per department

Terminology optimization

Glossary · continuous
  • Industry terms and proper nouns recognized first for accurate dictation
  • Share a team-wide glossary for consistent output
  • Keep improving as your business evolves

Privacy & internalization

Localized · auditable
  • Audio and transcripts stay fully local by default
  • Fine-grained access control and audit trails
  • Meet internal compliance and data-retention policies

Download TakeVox

macOS supports Apple Silicon (ARM) only — no Intel; pick the matching Windows or Linux architecture below.

macos

macOSARM only

Apple Silicon (ARM)
Requires macOS 13 or later
coming soon
Requires macOS 13 or later
windows

Windows

x64
coming soon
linux

Linux

x86_64
aarch64 planned
coming soon
aarch64 planned

Pro benefits

Subscribe in-app to unlock cloud transcription and meeting enhancements

Cloud transcription: longer audio, no model management Speaker separation for meeting minutes Sync across devices, resume anywhere Always up to date with the latest capabilities
App StorePayPalWeChat PayAlipay
Download

FAQ

Which operating systems are supported?
macOS (Apple Silicon/ARM), Windows 10/11 (x64) and major Linux distros (x86_64). macOS does not support Intel.
How large is the offline model?
The dictation model is a few hundred MB, downloaded automatically on first use, then fully offline.
Is my audio uploaded to a server?
Not by default. Local mode is fully offline; with cloud transcription your audio goes directly from you to the transcription provider — never through a TakeVox server.
How does cloud transcription work?
After subscribing in-app, use your own subscription key to talk to the cloud transcription service, which adds speaker separation and longer audio.
Do you support professional glossaries?
Yes. Add terms one at a time or bulk-import CSV/JSON industry lists in the glossary settings: misreadings are auto-corrected, preferred spellings are honored during polishing, and cloud transcription (own API) syncs them as hotwords so domain terms are recognized correctly the first time.
Where are transcripts and clipboard data stored?
Locally on your device. Clipboard history can be cleared in one click and specific apps ignored.
How do I pay for a subscription?
App Store, PayPal, WeChat and Alipay are supported, with renewal reminders.
Can I export meeting minutes?
Yes — Markdown export with timestamps, speaker labels and action items, plus jump-to-transcript links.
How do I get updates?
New versions are published on our website, and in-app update reminders keep you current.

Ready to turn speech into text?

Free to download. Local-first, private by default.

Sign in to manage your subscription, usage and account