Dictation glossary

Offline dictation

Offline dictation means speech-to-text that keeps working with no internet connection because the recognition model runs on your computer. You still need a microphone and enough memory and disk for the model. The first setup often downloads that model while you are online; afterward, audio does not need to leave the machine for transcription. Cloud dictation is the contrast: audio travels to a remote engine for each utterance.

In more detail

What is Offline dictation?

Offline dictation is a property of the engine, not of the hotkey. A local model files itself into application support or a cache, loads into RAM, and scores audio on the CPU, GPU, or a neural accelerator. Quality depends on which model you picked and whether your machine can run it in real time. Smaller models dictate faster and miss more; larger models are slower and hungrier. If the model never loaded, there is nothing to run offline — the download is a one-time online step, not a permanent tether.

People want offline mode for flights, trains, rural links, and for a privacy posture where dictation audio never traverses a network. It is not automatic anonymity: the text still exists on disk if the app keeps history, and the focused app may sync the resulting document elsewhere. Offline also does not mean you lack a cloud option in the same product. Some tools default to local recognition and offer a network engine as a switch for older hardware.

For writers who speak

Why it matters for dictation

Writers draft in places networks fail. If dictation dies the moment Wi-Fi does, it is a fair-weather tool. Offline engines also keep raw audio off of someone else's servers, which matters when you speak client names, medical notes, or source code. The trade is disk, RAM, and a first download — costs you pay once per machine instead of on every sentence.

In this product

How WhisperJot handles it

With the local engines — the default — WhisperJot works fully offline after a one-time model download. Jot Local downloads about 500 MB on a Mac (about 600 MB on disk on Windows and Linux). Jot Local Pro offers several models from about 80 MB up to 1.5 GB; its default on Mac is the 1.5 GB model. Jot Cloud is the opt-in alternative when you would rather skip that footprint; it needs a network, runs only while selected, and does not store audio on our servers.

Questions

Straight answers.

What is offline dictation?

Offline dictation is speaking to text without an internet connection, using a speech model that already lives on your device. After the model is installed, recognition does not need a server. You still need a microphone and enough hardware to run the model smoothly. Cloud dictation cannot do this, because each utterance is sent out for processing.

Do I need the internet the first time I use offline dictation?

Usually yes, once, to download the speech model. That file can be multiple gigabytes. After it is on disk, local dictation should start without a network. App updates, account login, or an optional cloud engine may still want connectivity. If a product claims offline use but never downloads a model, it is probably still calling a server.

Is offline dictation more private than cloud dictation?

It is more private with respect to the audio path: the waveform does not need to leave the machine for recognition. That is not the same as saying nothing is stored. The app may keep a local history, and the document you paste into may sync to a cloud office suite. Treat offline recognition as one scoped guarantee, not a blanket secrecy claim.

One hotkey, any focused app.

Private local transcription by default, with an optional opt-in cloud engine.