Blog · How-to

How to type by voice in any app on a Mac, with the speech recognised on the Mac itself

To type by voice anywhere on a Mac, you need a tool that puts the words at the cursor of whatever app is in front, rather than in a window of its own that you copy out of. Deixis does that: hold Ctrl+Option, speak, let go, and clean punctuated text lands at the caret — in an email, a document, a chat box, a browser form, a terminal or a code editor — with the speech recognised on the Mac itself. It is free, it does not ask you to sign in, and once the speech model is downloaded dictation works with the network switched off.

This post covers setting it up, the one gesture and the variations on it, what the Mac build does today and what it does not, what stays on your Mac, and how to get your own words right.

What you need, and setting it up

Deixis runs on a Mac with Apple silicon, on macOS 13 or later. Nothing else needs installing: no runtime and no redistributable.

  1. Download the disk image from the download link. It is notarized. Open it and drag Deixis to Applications.
  2. Fetch the speech model. The app does not carry the model's weights, so the first launch asks you to press Download once. Until that finishes there is no speech model, and dictation says so rather than failing quietly.
  3. Grant two permissions. The first time you dictate, macOS asks for Accessibility and microphone permission. Grant both in System Settings → Privacy & Security.

Updates are checked in the app, under Settings → Updates. The Mac has a release feed of its own, so it ships on its own schedule rather than waiting for Windows.

The one gesture: hold, speak, release

Hold Ctrl+Option, speak, release. The text lands at the caret of whatever has focus, and your clipboard is put back the way you left it. That is the product; everything else is a variation on the same chord. If the chord does not suit you, Settings → Hotkey captures a new one as you press it.

Double-tap the chord for free mode. It keeps recording until you double-tap again, so you can talk through a paragraph you have not finished thinking. The finished note is pasted at the caret, left on the clipboard and saved as a Markdown file.

Tap the chord during a free-mode session for the toolbar, a fan of actions at the cursor. Take the text you have selected, grab a region of the screen, or finish. What you said and what you pointed at come out as one note, interleaved — useful for a bug report, a review comment or instructions for somebody else.

Hold the chord and the mode palette appears above the listening pill, one chip per mode, each with its letter. There is more on the gestures page.

Screen grabs and screen recording

Two more things the Mac build does, both from the same chord.

  • Screen grab and annotation. The screen freezes, you drag a region, and you either quick-save it or open the editor: arrow, rectangle, ellipse, pen, highlighter, text, pixelate to redact, and numbered step badges. Every save keeps a sidecar file beside the image, so the grab stays re-editable rather than flattened. See screen capture.
  • Screen recording. It records the screen to a video file, and the same chord stops and saves it. While it runs, a small bar with a red dot, the elapsed time, Pause and Stop shows that it is rolling.

What the Mac build does not do

It is worth knowing before you download rather than after.

  • No meeting mode on the Mac. Recording and transcribing a meeting is on Windows and on the Android app, not on the Mac today. The platforms page keeps the current state of each build.
  • No translate on the Mac. Translating a selection is on Windows and Android.
  • English speech models. The two models that ship, base.en and small.en, are English.
  • Apple silicon only. There is no build for an Intel Mac.

What stays on your Mac, and what leaves

Speech recognition runs inside the app, on your Mac. The recording, the model and the transcript stay local, and nothing uploads the audio a dictation records, under any setting or on any plan. There is no telemetry.

A fresh install makes two calls on its own, and neither carries anything but the request: the update check at launch, and a check for an announcement from us when the Settings window opens, which has an off switch of its own. The speech model download happens once, when you press the button.

Cleanup can use a language model, and that route is off until you configure a destination. When cleanup does send something, it is text, never audio: the transcript, your vocabulary terms, the name of the app and field it is going into, and your previous sentence in that field — not the window title, which is where document names and customer names live. The one thing that uploads a file is Share a link, and only for the file you picked, when you pressed it. The full list, row by row, is on what leaves your machine.

Teach it your words

The dictionary is the biggest accuracy lever there is. Without any vocabulary, the standard model writes shadcn as "shadkin" and useEffect as "use effect", and one line of vocabulary fixes both. The app's Dictionary tab holds four things:

  • Words to expect — names and jargon the recogniser should listen for. Only the first 36 reach the speech model itself, because past that accuracy measurably degrades; the editor tells you when you are over.
  • Fixes — find-and-replace corrections applied after recognition, at word boundaries. These are unbounded, so once you are past the cap, put new corrections here.
  • Filler words — extra words to strip, beyond the built-in "um" and "uh".
  • Vocabulary libraries — five packs ship: software development (on by default), accounting and finance, business and product, medical, and legal. You can add your own.

One more lever needs no software at all: a headset beats a laptop's built-in microphone in exactly the conditions that break dictation, such as a room with other people in it. The dictionary page has the details.

Cleanup: rules, or a cloud model when you want one

Cleanup is the second pass that punctuates and structures what you said. With nothing configured it runs on rules on your Mac, and no text leaves. Settings → Cleanup decides when a language model is asked instead:

  • When it helps, the default. Rules handle the sentence locally unless you ask for a transform ("make this a bullet list") or dictate a long stretch with no punctuation.
  • Every time, for better prose on every sentence, at the cost of about a second on each sentence, and each sentence leaving the machine.
  • Never, which stops dictation calling a model at all.

With Deixis's own endpoint, which is what Pro pays for, requests route only to model endpoints that keep nothing and do not train on your text. The cleanup page covers the settings, and free and Pro the other ways to run the cloud features.

One thing cleanup cannot do is fix what the recogniser misheard. It rewrites prose, not hearing, so a misheard name stays misheard. Mishearings are the dictionary's job.

Before you ask

Questions people ask

How do I dictate into any app on a Mac?

Install Deixis, then hold Ctrl+Option, speak and let go. The text lands at the cursor of whatever app has focus — email, documents, chat, a browser form, a terminal — and your clipboard is put back afterwards.

Does Deixis dictation work offline on a Mac?

Yes. Speech recognition runs on your Mac, so after the one-time speech model download, dictation works with the network switched off. With no cleanup destination configured, cleanup runs on rules on your Mac and no text leaves it.

Does Deixis run on an Intel Mac?

No. The Mac build needs Apple silicon and macOS 13 or later.

Is Deixis free on a Mac?

Dictation, free mode, screen grabs and annotation, screen recording and the dictionary are free, with no account. Pro, at $49 a year or $10 a month, runs cloud cleanup for you, with no key to obtain.

Can Deixis record meetings on a Mac?

Not today. Meeting recording and transcription are in the Windows build and the Android app, and the Mac build does not have them.

Does my voice leave my Mac?

Not when you dictate: the audio of a dictation is recognised on your Mac and never uploaded, under any setting or on any plan. What you dictate leaves only as text, and only for the optional cloud features you configure. Deixis's one exception anywhere — a narrated screen recording you choose to share as a link — is described on what leaves your machine.

Keep reading

The one gesture

Hold Ctrl+Win (Ctrl+Option on a Mac) and speak, and the text lands at your caret. Plus every mode on the palette behind it, each on its own letter.

Where Deixis runs

Which parts of Deixis run on each of the four platforms, what state each build is in, and what is still being finished for meetings.

The dictionary, and getting your words right

How to stop Deixis mishearing you: vocabulary terms, find-and-replace fixes, filler words, the five packs, and the 36-term cap that decides which to use.

Here is every byte that leaves your machine, for each tool

Every byte Deixis sends off your machine in one table, and the same table filled in for other dictation tools from their own documentation, dated.

Is voice typing private? Eight things a dictation app can send, and how to check each one

Voice typing can send your audio, your words and even a screenshot. How to check what any dictation app sends, and exactly what Deixis sends.

How to transcribe a meeting without a bot joining the call

Record both sides of a call on your PC and transcribe it there, with no bot in the participant list and no audio uploaded. How it works, and its limits.

Two keycaps, Ctrl and Win, held down and lit teal from underneath.

Hold a key. Speak. Keep working.

Deixis is free, needs no account, and runs on your own device.

v1.2.2 · 64‑bit Windows 10/11 · macOS 13 on Apple silicon · 51 MB · nothing else to install

Everything this post says about Deixis is from its documentation as of the date above. Something out of date? Tell us.