← All posts

On-device vs cloud dictation — why it matters for your Mac

Most dictation apps you can install today are, underneath, a microphone pointed at someone else’s data centre. You speak, your audio is uploaded, a model on a server transcribes it, and the text comes back. It works — but you pay for it three times: with a monthly subscription, with a wait after every sentence, and with a copy of your voice you no longer control.

On-device dictation flips that. The model runs on your Mac’s Neural Engine. Your audio never leaves the machine.

Privacy: architecture, not policy

A cloud app can promise it deletes your audio. On-device dictation doesn’t need the promise — there’s no server to send anything to. Akova holds audio in memory just long enough to turn it into text, then it’s gone. You can dictate on a plane, on a train with no signal, or into a document you’d never paste into a web form.

Speed: no round trip

Every cloud transcription is a network round trip: upload, queue, process, download. On-device, the Neural Engine transcribes as you speak. There’s no “processing…” spinner between thoughts.

Cost: own it, don’t rent it

Cloud transcription costs the provider money for every second of audio, so the business model is a subscription — $12–$15 every month, forever. On-device inference costs nothing to run, which is why Akova can be a one-time purchase. Buy it once, own it.

The catch (and why it’s fine now)

On-device transcription used to mean worse accuracy. That gap has closed: Whisper-class models running on Apple Silicon are genuinely good, and a local clean-up pass turns rough speech into polished writing. The only thing you give up is the monthly bill.

Ready to try it? Download Akova free — no account, no card.