| How it works |
Hold a key, speak, release; the verbatim transcript pastes at the cursor. |
Press the Dictation shortcut to start, stop manually; auto-stops after 30 s of silence. |
Hotkey recording; optional AI modes reshape the output before insertion. |
Hold hotkey; AI-formatted text appears. |
Hotkey dictation with optional AI enhancement. |
| Price |
Free. |
Included with macOS. |
Free tier; Pro starts at $8.49/month. |
Free tier of 2,000 words/week on desktop; Pro is $15/month, or $12/month billed annually. |
From $25 one-time for one Mac; free to build from source. |
| Source |
Open source, MIT. |
Proprietary, built into macOS. |
Proprietary. |
Proprietary. |
Open source, GPL-3.0 per its README. |
| Where audio is processed |
On this Mac only; no cloud transcription path exists in the app. |
On-device for supported languages; otherwise sent to Apple servers. Keyboard settings shows which applies. |
Your choice per model: fully local models or cloud services. |
In the cloud — “Transcription always happens in the cloud” per its privacy page. |
On-device by default; optional cloud providers. |
| Account required |
No. |
No. |
No sign-in documented; Pro activates with a license key. |
Yes. |
No; the paid binary activates with a license key. |
| Published latency evidence |
Model-only p50 of 92–152 ms across the published clips; the reproducible benchmark does not time clipboard or paste work. |
Not published. |
Not published. |
Its engineering blog describes a ~700 ms end-to-end target, which is not directly comparable with Presspeech's model-only measurement. |
Not published. |
| App size |
8.4 MB release zip, plus about 500-600 MB for the local model cache. |
Built in; on-device languages need a model download, size not published. |
Not published; local models are 75 MB–3 GB. |
Not published; docs ask for ~500 MB free storage. |
47.8 MiB download (v2.13 DMG); models extra. |
| Idle footprint |
~80 MB RAM, 0% CPU between dictations. |
Not published. |
Not published. |
Not published; 8 GB+ RAM recommended. |
Not published; 8 GB RAM recommended. |
| Languages |
25 European languages, auto-detected; Settings can pin one of 18. |
63 Dictation locales, 48 of them with on-device processing. |
100+ via cloud models; local model counts vary. |
100+. |
No official count; its Whisper models cover about 100. |
| Platforms |
Released on Apple Silicon Macs, macOS 14+; a separate Windows preview is available. |
Every Mac. |
macOS 13.3+, Windows, iOS; local models work best on Apple Silicon. |
macOS 11+ including Intel, Windows, iPhone, Android. |
Apple Silicon Macs, macOS 14.4+. |
| Beyond verbatim dictation |
Deterministic local corrections and shortcuts, opt-in spoken formatting, and filler-word removal; no AI rewriting. |
Auto-punctuation, spoken punctuation and emoji commands. |
AI modes, meeting transcription, file transcription, translation. |
AI formatting, tone styles, command mode, cross-device dictionary. |
AI enhancement modes, per-app Power Mode profiles, file transcription. |