Dictation in QUILL and QUILL Lite
Press Ctrl+F11, talk, and pause. Each time you pause, what you said is written into the document at the cursor, a soft tone tells you it arrived, and the words are read back so you know they are right. Then it keeps listening. The speech recognition ships inside both programs: there is nothing to download, no account to make, nothing you say leaves your computer, and no recording is kept anywhere.
This page covers all of it: every key, every phrase dictation understands, every setting, the four speech engines, and what to do when it mishears you. It covers both programs, and tells you wherever they are different.
Your first dictated sentence
This works the same in QUILL and in QUILL Lite. Nothing needs setting up first, and you do not need to say any punctuation. Before you start, the way to stop is Ctrl+F11 again, or saying "stop dictation", or Escape to throw away the phrase being heard. It helps to know that first if you cannot see whether the microphone is on.
- Put the cursor where you want the words. If you select some text first, the first phrase you say replaces it.
- Press Ctrl+F11. The first time after opening the program it takes a second or two while the speech engine loads. Then you hear two rising tones and "Dictation on".
- Say a sentence the way you would say it to a person, and pause.
- A moment later the sentence is in the document, you hear one short soft tone, and the words are read back to you.
- Keep going. There is nothing to press between sentences.
- When you are done, press Ctrl+F11 again, or say "stop dictation". You hear two falling tones and "Dictation off".
A phrase is written once you have been quiet for a little under a second, so a breath in the middle of a sentence does not cut it in two. If a pause does split a sentence and the next part starts with a word like and, but, which or to, the full stop the pause put in is taken back out and the sentence carries on.
Every phrase is one step for Ctrl+Z, punctuation and capitals included, even a phrase that replaced a selection.
Holding Ctrl+F11 to talk
Press Ctrl+F11 to start dictating and press it again to stop. That is all most people ever need, and it is how dictation starts out.
If you prefer to hold the key while you talk, like the button on a walkie-talkie, turn on Hold the dictation key to talk; a quick press still turns it on and off in More Dictation Settings. Then:
- Put the cursor where you want the words.
- Press and hold Ctrl+F11. You hear the two rising tones and "Dictation on" straight away.
- Keep holding the keys, and say what you want to say.
- Let go. Dictation writes your last phrase, then you hear the two falling tones and "Dictation off".
With that turned on, a quick press still works just the same: dictation turns on and stays on until you press Ctrl+F11 again. Half a second or more is a hold. Stopping never cuts you off: if you press or let go while you are still finishing a sentence, dictation writes that phrase first.
What you hear
Four moments, each with its own tone and its own words. Both halves can be changed, and two of them can never be silenced.
| When | Tones | Words |
|---|---|---|
| Dictation starts | two rising tones | "Dictation on" |
| A phrase is written | one short, soft tone | the words that were written |
| Dictation stops | two falling tones | "Dictation off" |
| Something goes wrong | a low double tone | what happened, and what to do about it |
Every row can be changed in Dictation Settings, and every tone can be changed or silenced in Tools, Sound Scheme. Two things are never silenced: a failure is always spoken, and so is the answer to a spoken command, because both answer a question.
The soft tone plays only once the words are really in the document, so hearing it means they arrived. If you have the words read back, they follow a quarter of a second after they are written, so your screen reader's own reaction to the new text does not cut them off.
Use headphones if the read-back is on. Through speakers, the microphone can hear the read-back and write it down a second time. With speakers, set "After each phrase is written, give me" to a sound only.
How to tell it is listening, without sound and without sight
Every outcome above is also written down, so none of this depends on hearing a tone. Two ways, and both reach a braille display. The status bar has a Dictation part that says Off, Listening, Hearing you, Writing, Spelling or Waiting for wake phrase, and F6 moves focus there so you can read it cell by cell. And Tools, Dictation, Dictation On is a check box in the menu, so opening that menu announces it as checked or unchecked.
The same holds for failures: a failure is always spoken and written to the status bar, whatever the sound settings say, and it is never reduced to a tone. If you would rather have no tones at all, silence them in Tools, Sound Scheme and nothing is lost.
Hearing the punctuation in the read-back
When a phrase is read back to you, its punctuation is read too, using the same names you would say to dictate it: "Hello comma world period". Quotes, question marks and new paragraphs are said the same way: "open quote", "question mark", "new paragraph". You hear them whatever your screen reader's punctuation level is, so you know exactly what was written without stopping to check.
A few things stay quiet, because saying them would only get in the way: an apostrophe or a hyphen inside a word, and the full stop or comma inside a number such as 3.5 or 1,000. A web address is said the way you would dictate it, so example.com is read as "example dot com". In Spanish you hear the Spanish names, such as "coma" and "punto".
The status bar and a braille display show the real characters, as in "Dictated: Hello, world." If you would rather hear just the words, uncheck Say punctuation marks in the read-back in More Dictation Settings. It is checked to begin with.
The two kinds of dictation
QUILL has two, and they work quite differently. QUILL Lite has the first one.
Live Dictation: writes as you go
Ctrl+F11, in both products. Each phrase is written the moment you pause, so you hear it land and can correct it by voice straight away. This is the one described above and for the rest of this page unless it says otherwise. In QUILL it is under Tools, Speech, Live Dictation; in QUILL Lite it is Tools, Dictation.
Locked Dictation: records a passage, then writes it out
Ctrl+F9, QUILL only. Press it once and speak as long as you like, across sentences and paragraphs; press it again and QUILL transcribes the whole passage and inserts it as one undoable edit. It uses QUILL's on-device Whisper engine rather than the live engines, so nothing is uploaded here either. Use it when you would rather talk without interruption and read the result afterwards.
- While recording, Escape stops and keeps what you said for transcription; Shift+Escape cancels and discards it.
- A session stops by itself after five minutes, and stops and preserves your audio if QUILL loses focus.
- Audio is saved to a recovery folder before transcription starts, so a crash cannot lose a passage. Tools, Speech, Locked Dictation, Dictation History and Review lists every recording whose transcript was never inserted, so you can insert, copy or discard it. If anything is waiting at startup, QUILL says so.
Dictate (Offline): a simple on and off switch
QUILL only, on Ctrl+Shift+Grave, D. Start it, talk, stop it: QUILL transcribes in the background and inserts the text at the cursor as one edit, with the word count in the status bar. It shares its engine and microphone with Locked Dictation.
Live Dictation and Locked Dictation each remember their own microphone choice.
Every dictation key
These are the keys it starts with, and you can change any of them in QUILL Lite's Keyboard Manager (Ctrl+Alt+Shift+R) or QUILL's Settings, Keyboard. Each row starts with what you want to do, then the key to press.
Keys both products share
These keys are the same in QUILL and QUILL Lite, so a key you learn in one works in the other.
| To do this | Press | Notes |
|---|---|---|
| Start or stop dictation | Ctrl+F11 | Press once to start and again to stop, or hold it to talk if you turn that on. Also works in the Find box, either Replace box, the AI pad's question box and the AI Conversation window's message box |
| Throw away the phrase being heard | Escape | Only while the status says "hearing you"; otherwise Escape does whatever it always did. While the AI's reply is being read, Escape listens at once |
| Open Dictation Settings | Alt+Shift+F6 | Every option on this page lives there |
| See the last twenty phrases | Shift+F11 | Insert Again writes one back as a single undo step; Copy puts it on the clipboard |
| Teach dictation your words | Alt+Shift+F10 | Words, phrases and corrections, in a window |
| Turn a recording into text | Shift+F5 | Transcribe a Recording, in the background with any speech model you have. Press it again to see how far it has got, or to stop it |
| Tidy dictated text with AI | Ctrl+F3 | Needs a ChatGPT subscription or your own AI key (OpenAI or Google Gemini) |
| Take back the last phrase | Ctrl+Z | One phrase per press, whatever it replaced |
| Switch between English and Spanish | Ctrl+Shift+F11 | Switch Dictation Language. Works while you are dictating, and the choice is kept for next time |
| Start or stop a live transcript | Ctrl+Alt+Shift+PageDown | Writes everything you say into a new document of its own |
| Describe this document for dictation | Ctrl+Alt+Shift+PageUp | Dictation Context for This Document, used by OpenAI dictation and Tidy Dictated Text |
| Hear what dictation is doing | Alt+F9 | Dictation Status. Says it once and changes nothing |
Keys only QUILL has
Locked Dictation and the offline switch. QUILL Lite keeps things simple and has Live Dictation only.
| To do this | Press | Notes |
|---|---|---|
| Start or finish Locked Dictation | Ctrl+F9 | Records a whole passage, then transcribes and inserts it as one edit |
| Pause or resume the recording | Ctrl+Shift+F9 | Nothing is lost while paused |
| Hear the current state | Alt+F9 | While Locked Dictation is recording, speaks what it is doing. The rest of the time the same key is Dictation Status, as in QUILL Lite |
| Stop and keep what you said | Escape | While recording. The passage is still transcribed |
| Cancel and discard what you said | Shift+Escape | While recording. Nothing is transcribed and nothing is inserted |
| Dictate with the simple toggle | Ctrl+Shift+Grave, D | Dictate (Offline): start, talk, stop, and the text arrives at the cursor |
Ctrl+Shift+Grave, D is two presses, not four keys at once: press the QUILL Key (Ctrl+Shift+Grave), let it go, then press D. QUILL Lite does not use two-step keys like this. To see what else is different, read how QUILL and QUILL Lite compare.
The speech engines
Four choices for Live Dictation, in Dictation Settings. The first two ship inside both programs, so they work with no download and no connection. Larger models you choose to download join the list (Better accuracy: optional speech models).
| Engine | What it is like | Punctuates by itself |
|---|---|---|
| Moonshine | The one it starts with. Fast even on a modest computer, and the most accurate of the field. Englishbuilt in | Yes |
| Whisper | A little slower. Worth trying if Moonshine often mishears your voice or your microphone. Englishbuilt in | Yes |
| Windows speech recognition | Windows' own recogniser. Needs nothing extra and can use any speech language installed in Windows, but mishears an untrained voice oftensay every mark yourself | No |
| Windows voice typing | Hands over to the panel Windows+H opens. Windows does the recognising and the typing, so none of dictation's commands, tones, read-back or wake phrase applyfor people who already prefer it | Windows decides |
| OpenAI | Only with your own OpenAI key, and only if you choose it and agree: your speech is sent to OpenAI. Very accurate, and can show your words as you speakOpenAI dictation, with your own key | Yes |
QUILL's Locked Dictation uses a different set: on-device Whisper, or Parakeet 3 once you install it from Tools, Speech, Manage Speech Models. Parakeet is worth having because, when you are silent, Whisper sometimes makes up words, like a stray "thank you". Parakeet only writes what it really hears. Once you install it, QUILL uses it unless you have picked a different engine yourself. If a recording has no speech in it at all, QUILL tells you so instead of guessing.
How the default was chosen
We tried eight engines on ten everyday sentences each, on 25 September 2026. We wanted one that keeps up with you on an older Windows 10 computer, puts in its own punctuation, and is small enough to come inside the program.
| Engine | Word errors | Punctuation, found and right | Seconds of computing per second of speech | Size shipped |
|---|---|---|---|---|
| Windows speech, SAPI | 9 per cent | none, and none | 0.09 | nothing to ship |
| Vosk small | 4 per cent | none, and none | 0.32 | about 45 MB |
| Vosk small, plus a punctuation model | 4 per cent | 7 per cent, and 10 per cent | 0.33 | about 52 MB |
| Moonshine tiny, the default | 2 per cent | 93 per cent, and 87 per cent | 0.04 | about 124 MB |
| Moonshine base | 2 per cent | 50 per cent, and 88 per cent | 0.06 | about 287 MB |
| Whisper tiny English, the alternative | 3 per cent | 86 per cent, and 86 per cent | 0.15 | about 104 MB |
| Whisper base English | 2 per cent | 93 per cent, and 87 per cent | 0.28 | about 161 MB |
| Parakeet 3 | 2 per cent | 93 per cent, and 87 per cent | 0.21 | about 670 MB |
Anything under 1.0 in the speed column keeps up with you as you talk. Moonshine tiny was the most accurate and by far the fastest, so it works well even on an older computer. Whisper tiny is there for a voice or microphone that Moonshine struggles with. Parakeet 3 is very good but five times the size, so it is an optional download in QUILL rather than something packed into QUILL Lite.
One thing to keep in mind. For this test, Windows' own computer voice read the sentences. That is easier to understand than a real person, so every engine will make more mistakes with your voice than the table shows. Windows speech recognition in particular does much worse with a real voice. We will update these numbers after testing with real voices.
Better accuracy: optional speech models
Dictation works the moment you install. If you want more accuracy, you can download a larger model: the same local models VS Code offers, plus the rest of the Whisper family. Free, optional, and run on your computer's processor; no graphics card needed.
Your voice never leaves the computer. The only thing that comes over the internet is the model itself, once, when you ask for it. Nothing downloads on its own, and the built-in Moonshine stays your engine until you choose another. Both QUILL and QUILL Lite have this, in the same place, and share the models on one computer.
Downloading one
- Open Dictation Settings (Alt+Shift+F6) and press Better Accuracy: Speech Models... (Alt+B).
- Arrow through the list. Each row says what the model is good for, its size, and whether it is here. Details says what it is better at, its languages, download and disk size, which computers suit it, whether this computer should keep up, its published accuracy, its licence and where it will be saved.
- Press Download.... A question names where it comes from, its size, its licence and the folder. Answer Yes. On a metered connection you are asked about that first, and if the drive is too full you are told before anything starts.
- You hear "Downloading", then every quarter of the way. Cancel Download stops it and keeps what arrived, so Download carries on later. Every file is checked against a recorded checksum before it is used.
- Press Use for Dictation, then OK in Dictation Settings. Or choose it in the Speech engine list, where downloaded models show "(downloaded)".
Where they go: %LOCALAPPDATA%\QuillVille\Dictation\models,
one folder for both programs. In a portable copy, models are saved inside the portable
folder, so everything travels together. Remove in the same window
frees the space. If a downloaded model goes missing, dictation says so in one sentence
and carries on with the built-in engine; it never switches silently.
Will my computer keep up? When the window opens it spends about two seconds timing the built-in Moonshine, then tells you for each model whether this computer should keep up comfortably, keep up but leave your screen reader less room, or "may lag behind your speech on this computer". You can still choose any of them.
Words appear when you pause, with every engine. Nemotron is built to recognise while you speak; showing its words as you talk is planned.
The models, in our suggested order
The order is our suggestion, not a measurement: VS Code's own models first, then the rest of Whisper, smallest first.
| Model | Good for | Download | Languages |
|---|---|---|---|
| NVIDIA Nemotron 3.5 ASR Streaming 0.6B | Suggested download. Excellent for live dictation; VS Code's default | 682 MB | English, Spanish |
| NVIDIA Parakeet Unified 0.6B | May give slightly better final text in some conditions | 663 MB | English |
| NVIDIA Parakeet TDT 0.6B v3 | Excellent recognition of each finished phrase | 670 MB | English, Spanish |
| Whisper small | A VS Code alternative: good | 375 MB | English, Spanish |
| Whisper base | A VS Code alternative: moderate | 161 MB | English, Spanish |
| Whisper tiny | Built in already. Fastest and lightest, weakest accuracy | built in | English, Spanish |
| Whisper base.en | Light and quick, a little better in English than base | 161 MB | English |
| Whisper small.en | Good English accuracy for a capable computer | 376 MB | English |
| Distil-Whisper small.en | Small's accuracy, made smaller and faster | 299 MB | English |
| Distil-Whisper medium.en | Close to medium's accuracy at a fraction of the work | 573 MB | English |
| Whisper medium.en | Best for accuracy when you can wait | 946 MB | English |
| Distil-Whisper large-v3 | Near large-v3 accuracy in English, much faster | 984 MB | English |
| Whisper medium | Best for accuracy when you can wait | 946 MB | English, Spanish |
| Whisper large-v3-turbo | Large-v3 quality, much faster than large-v3 | 1.0 GB | English, Spanish |
| Whisper large-v3 | The most accurate Whisper; very slow on most computers | 1.8 GB | English, Spanish |
| Moonshine base | A light step up from the built-in Moonshine | 141 MB | English |
English only, or many languages? Whisper models ending in ".en", and Distil-Whisper, know only English and are a little more accurate in English than the many-language model of the same size. The many-language ones also dictate Spanish.
How accurate they are
These are the figures each model's makers publish. A word error rate is the share of words a model gets wrong on a recorded test set; lower is better. Test sets differ, so compare numbers only within one line.
- Nemotron 3.5 ASR Streaming: 7.91 per cent on FLEURS English and 4.11 per cent on FLEURS Spanish, at the chunk size this download uses (NVIDIA's model card).
- Parakeet Unified: 1.63 per cent on LibriSpeech test-clean, 3.11 on test-other (model card).
- Parakeet TDT v3: 4.85 per cent on FLEURS English, 3.45 on FLEURS Spanish (model card).
- Whisper on LibriSpeech test-clean: tiny 7.6, base 5.0, small 3.4, medium 2.9 per cent; tiny.en 5.6, base.en 4.2, small.en 3.1, medium.en 3.1. On Spanish (Common Voice 9): tiny 30.3, base 19.6, small 10.3, medium 6.9 (OpenAI's Whisper paper, Appendix D).
- Distil-Whisper and large-v3, average over short recordings the models were not trained on: large-v3 8.4, distil-large-v3 9.7, distil-medium.en 11.1, distil-small.en 12.1 per cent (Distil-Whisper model card). OpenAI publishes no figure for large-v3-turbo; it says turbo is much faster than large-v3 with a small drop in quality.
- Moonshine: base 10.07 per cent against tiny's 12.66, averaged over the Open ASR Leaderboard's English sets (Moonshine AI).
In plain words: the bigger models cope better with accents, fast speech, noisy rooms, names and long sentences; every one puts in punctuation and capitals; and for Spanish any many-language model beats the built-in tiny by a wide margin.
What we measured on our test computer on 5 October 2026: ten sentences and two spoken commands read by Windows' own computer voice, one processor thread, a 12-core desktop with 8 GB of memory. A computer voice is easy to understand, so the models scored about the same on accuracy, and all of them heard both commands. The cost is what differs:
| Model | Word errors | Seconds of computing per second of speech | Text after you pause | Most memory | Load time |
|---|---|---|---|---|---|
| Moonshine tiny (built in) | 3.4% | 0.05 | 0.2 s | 210 MB | 2.2 s |
| Whisper tiny.en (built in) | 3.4% | 0.20 | 0.7 s | 303 MB | 1.3 s |
| Moonshine base | 3.4% | 0.08 | 0.3 s | 365 MB | 2.0 s |
| Whisper base | 5.1% | 0.37 | 1.4 s | 417 MB | 1.5 s |
| Parakeet Unified 0.6B | 3.4% | 0.33 | 1.2 s | 844 MB | 5.2 s |
| Parakeet TDT 0.6B v3 | 4.2% | 0.34 | 1.2 s | 805 MB | 5.1 s |
| Nemotron 3.5 ASR Streaming | 2.5% | 0.40 | 1.5 s | about 790 MB | 3.2 s |
| Whisper small | 5.1% | 1.29 | 4.7 s | 859 MB | 4.4 s |
| Whisper medium | 3.4% | 3.04 | 11.2 s | 2.1 GB | 9.6 s |
| Whisper large-v3-turbo | 3.4% | 4.20 | 15.4 s | 1.4 GB | 5.3 s |
Dictation uses two threads, and 0.25 is the budget the built-in engines meet so your screen reader always has room. The built-in model is still the better choice on an older or dual-core computer, on battery, or when your screen reader must never slow down.
Nemotron or Parakeet: which should I try?
Both are NVIDIA models of about 0.6 billion parameters, both are free, and both run here on the processor alone. The difference is when they listen.
- Nemotron 3.5 ASR Streaming is built for live dictation: it recognises while you speak, in short slices, for low delay, and writes punctuation and capitals itself, in English, Spanish and many other languages. Its words appear in the status bar and on braille while you talk (Seeing the words while you speak), and each phrase is written when you pause. It decides whether a sentence was a question once it hears the next one, so a full stop it wrote can become a question mark a moment later. About 682 MB.
- Parakeet recognises each finished phrase after you pause. Nothing appears until you stop, but having the whole phrase can give slightly more accurate final text in some conditions, question marks included. Parakeet TDT 0.6B v3 knows English, Spanish and 23 other European languages (670 MB). Parakeet Unified 0.6B is newer and English only; NVIDIA trained it to work both while you speak and on a finished phrase, and QUILL uses it on the finished phrase (663 MB).
On our test computer, on one thread, Parakeet Unified needed 0.33 seconds of computing per second of speech, Parakeet TDT 0.34 and Nemotron 0.40; each used about 800 MB of memory. Nemotron, fed at live pace, had its first words about 1.7 seconds into each sentence.
Rule of thumb: for writing as you go, try Nemotron; for dictating a paragraph and then checking it, try Parakeet.
How to tell which is more accurate for you
- Download one model, choose it, and dictate as you normally do for a day.
- Note each phrase that comes out wrong. Recent Phrases (Shift+F11) shows what was written.
- Switch to another model in the Speech engine list and do the same.
- Keep the one that made fewer mistakes, and remove the other.
For an exact answer, QUILL's benchmark measures any model on your own recordings: ten
16 kHz mono WAV files and a transcripts.txt of what you said, then
python scripts/dictbench.py folder --model nemotron in a copy of QUILL's
source. It reports the share of words wrong, the marks written, the speed, the memory
and the load time.
Everything you can say
There are 78 punctuation marks and layout phrases, and 41 spoken commands. This list comes straight from dictation itself, so it is always complete. While you are dictating, say "what can I say" to hear the same list in a window.
What you can say, and what it writes
Punctuation, brackets, symbols and layout. Each table lists the words to say and what appears in your document.
Punctuation
Say any of these anywhere in a phrase. Spaces go where they belong: none before a comma, one after a full stop.
| Say | What it writes |
|---|---|
| period | a full stop . |
| full stop | a full stop . |
| comma | a comma , |
| question mark | a question mark ? |
| exclamation point | an exclamation mark ! |
| exclamation mark | an exclamation mark ! |
| colon | a colon : |
| semicolon | a semicolon ; |
| ellipsis | three dots, an ellipsis ... |
| open quote | a double quotation mark " |
| begin quote | a double quotation mark " |
| close quote | a double quotation mark " |
| end quote | a double quotation mark " |
| open single quote | a single quotation mark, also the apostrophe ' |
| close single quote | a single quotation mark, also the apostrophe ' |
| apostrophe | a single quotation mark, also the apostrophe ' |
Brackets
Both halves are separate phrases, so you can say one without the other.
| Say | What it writes |
|---|---|
| open parenthesis | an opening round bracket ( |
| open paren | an opening round bracket ( |
| left parenthesis | an opening round bracket ( |
| close parenthesis | a closing round bracket ) |
| close paren | a closing round bracket ) |
| right parenthesis | a closing round bracket ) |
| open bracket | an opening square bracket [ |
| left bracket | an opening square bracket [ |
| close bracket | a closing square bracket ] |
| right bracket | a closing square bracket ] |
| open brace | an opening curly brace { |
| close brace | a closing curly brace } |
| open angle bracket | an opening angle bracket, the less-than sign < |
| close angle bracket | a closing angle bracket, the greater-than sign > |
Dashes and joining
What "dash" writes is up to you, in Dictation Settings. This row shows what it writes until you change it.
| Say | What it writes |
|---|---|
| hyphen | a hyphen - |
| dash | a dash, as chosen in Dictation Settings — |
| slash | a forward slash / |
| forward slash | a forward slash / |
| backslash | a backslash \ |
| underscore | an underscore _ |
Symbols
The symbols people dictate most often in ordinary prose and in code.
| Say | What it writes |
|---|---|
| at sign | an at sign @ |
| hash sign | a hash, the number sign # |
| number sign | a hash, the number sign # |
| dollar sign | a dollar sign $ |
| percent sign | a percent sign % |
| ampersand | an ampersand & |
| asterisk | an asterisk * |
| plus sign | a plus sign + |
| minus sign | a hyphen, the same character that "hyphen" writes - |
| equals sign | an equals sign = |
Lines and layout
These change where the next words go rather than adding a character.
| Say | What it writes |
|---|---|
| new line | a line break, ending the line |
| newline | a line break, ending the line |
| new paragraph | a blank line, starting a new paragraph |
| tab key | a tab |
| press tab | a tab |
| tab | a tab |
Markdown and code
For Markdown and code. A backtick opens or closes depending on how many are already on the line, and a code fence goes on a line of its own.
| Say | What it writes |
|---|---|
| backtick | a backtick ` |
| back quote | a backtick ` |
| triple backtick | a code fence, three backticks on a line of their own ``` |
| code fence | a code fence, three backticks on a line of their own ``` |
| tilde | a tilde ~ |
| vertical bar | a vertical bar | |
| pipe symbol | a vertical bar | |
| caret | a caret ^ |
| greater than sign | a closing angle bracket, the greater-than sign > |
| less than sign | an opening angle bracket, the less-than sign < |
Starting a line (say these first)
These work only at the start of a phrase, so a sentence that happens to contain the words stays words. If the cursor is not at the start of a line, a new line is started first.
| Say | What it writes |
|---|---|
| bullet | a hyphen and a space, starting a bullet - |
| list item | a hyphen and a space, starting a bullet - |
| numbered item | a 1, a full stop and a space, starting a numbered item 1. |
| block quote | a greater-than sign and a space, starting a block quote > |
| heading one | one hash sign and a space, starting a level 1 heading # |
| heading 1 | one hash sign and a space, starting a level 1 heading # |
| heading two | two hash signs and a space, starting a level 2 heading ## |
| heading 2 | two hash signs and a space, starting a level 2 heading ## |
| heading three | three hash signs and a space, starting a level 3 heading ### |
| heading 3 | three hash signs and a space, starting a level 3 heading ### |
| heading four | four hash signs and a space, starting a level 4 heading #### |
| heading 4 | four hash signs and a space, starting a level 4 heading #### |
| heading five | five hash signs and a space, starting a level 5 heading ##### |
| heading 5 | five hash signs and a space, starting a level 5 heading ##### |
| heading six | six hash signs and a space, starting a level 6 heading ###### |
| heading 6 | six hash signs and a space, starting a level 6 heading ###### |
Commands dictation obeys
A command works only when it is the whole phrase, said on its own after a pause. "Delete that line of text" in the middle of a longer sentence is just words, and is written as words. Say "literal" before any phrase on this page to write it out instead of acting on it.
Correcting
Each of these works only as a whole phrase, said on its own after a pause, and only while the phrase it changes is still exactly as dictation wrote it.
| Say, on its own | What happens |
|---|---|
| scratch that or delete that | Removes the phrase you dictated last. Say it again to remove the one before. A phrase you have typed into since is left alone. |
| undo that or undo or undo last | The same as pressing Ctrl+Z. |
| select that | Selects the phrase you dictated last, so you can change or format it with the keyboard. The next phrase you say replaces it. |
| capitalize that or cap that | Gives every word of the last phrase a capital: meeting notes becomes Meeting Notes. |
| all caps that or uppercase that | Puts the last phrase in capitals. |
| no caps that or lowercase that | Puts the last phrase in small letters. |
| delete word or delete last word | Deletes the word just before the cursor. |
| delete sentence or delete last sentence | Deletes from the start of the sentence the cursor is in up to the cursor. |
| read that or repeat that | Reads the last phrase aloud again. |
| correct that | Reads the other things the speech engine thought you said, numbered, when it offers them (Windows speech recognition does). Then say choose and a number. |
| choose one or choose 1 | After correct that, puts the first of the other guesses in place of the last phrase. |
| choose two or choose 2 | The same, with the second guess. |
| choose three or choose 3 | The same, with the third guess. |
| spell that | Selects the phrase you dictated last and listens for it spelled out. Say the letters and they replace it. The fix for a name the engine keeps getting wrong. |
Moving the cursor
Said on their own, after a pause.
| Say, on its own | What happens |
|---|---|
| go to beginning of line or go to start of line | Moves the cursor to the start of the line. |
| go to end of line | Moves the cursor to the end of the line. |
| go to top or go to start of document or go to beginning of document | Moves the cursor to the start of the document. |
| go to bottom or go to end of document | Moves the cursor to the end of the document. |
| go to start of paragraph or go to beginning of paragraph | Moves the cursor to the start of the paragraph. |
| go to end of paragraph | Moves the cursor to the end of the paragraph. |
Dictation itself
Said on their own, after a pause.
| Say, on its own | What happens |
|---|---|
| start spelling or spell mode or spelling mode | Starts spelling mode: every phrase is read as letters until you say stop spelling. See Spelling below. |
| stop spelling or end spelling or spelling off | Leaves spelling mode. |
| what can i say or show commands or dictation commands | Opens this list. |
| stop dictation or stop dictating or stop listening | Stops dictation. With a wake phrase set, dictation goes back to waiting for it. |
| switch to spanish or spanish dictation | Dictate in Spanish from now on, the same as Switch Dictation Language. In Spanish, cambiar a inglés or dictado en inglés comes back. |
| switch to english or english dictation | Dictate in English from now on. |
Selecting
Say the words you want, and dictation finds the nearest place they appear. If they are not there, your phrase is written as ordinary text and you are told.
| Say, on its own | What happens |
|---|---|
| select sentence or select this sentence | Selects the sentence the cursor is in. |
| select line or select this line | Selects the line the cursor is on. |
| select paragraph or select this paragraph | Selects the paragraph the cursor is in. |
| select again or select next | After select, go to or correct with words, moves on to the next place those words appear. |
| select previous | The same, going back to the place before. |
Capitals and spacing
Each one lasts until you turn it off or stop dictation, and the status bar shows it while it is on.
| Say, on its own | What happens |
|---|---|
| caps on or capitals on | Every word you say starts with a capital until you say caps off. Handy for titles and names. |
| caps off or capitals off | Words go back to ordinary capitals. |
| all caps on | Everything you say is written in capitals until you say all caps off. |
| all caps off | Back to ordinary capitals. |
| no space on | Words are written joined together with no spaces until you say no space off. Handy for web addresses and file names. |
| no space off | Spaces between words come back. |
Clips, snippets and copying
A snippet or a clip goes in as one phrase, so one Ctrl+Z or "scratch that" takes it back out.
| Say, on its own | What happens |
|---|---|
| copy all or copy everything or copy document | Copies the whole document, the same as Copy All on the Edit menu. Nothing is selected. |
| copy that | Copies the phrase you dictated last. |
| show clips or open copy tray or show copy tray | Opens the Copy Tray, where you choose a slot to paste. |
| show snippets | Opens the list of snippets to choose from. |
Spelling a word out
- Say "start spelling", then letters. Everything you say is written as letters until you say "stop spelling".
- Say "capital" before a letter for a capital, and "space" for a space.
- Letter names work ("bee", "see"), but the phonetic alphabet works much better: alpha, bravo, charlie, delta, echo, foxtrot, golf, hotel, india, juliet, kilo, lima, mike, november, oscar, papa, quebec, romeo, sierra, tango, uniform, victor, whiskey, x-ray, yankee, zulu.
- Numbers are written as digits: zero, one, two, three, four, five, six, seven, eight, nine.
- Punctuation goes in with no spaces while you spell, so an email address works: "jay dot smith at sign example dot com". "Dot" and "point" are a full stop only while spelling.
- Say "all caps" for capitals until you say "no caps" or the phrase ends.
- To spell just one word, say "spell" and the letters: "spell bravo alpha delta" writes bad. A sentence that only starts with the word "spell" is still written as words.
- Say "spell that" to select the phrase you just said. The next phrase is spelled and takes its place.
Starting and stopping without the keyboard
- With the wake phrase switched on in Dictation Settings, say "Quill dictate" to start dictation without touching a key. Anything you say after it in the same breath is written.
- Say "stop dictation" on its own, after a pause, to stop. With the wake phrase on, stopping goes back to waiting for the wake phrase, so you need never touch the keyboard at all.
- You choose both phrases yourself, and each needs at least two words, so ordinary talk cannot start or stop dictation by accident.
Correct that: choosing another guess
Sometimes the engine had a second idea about what you said. Say "correct that" straight after the phrase, and you hear up to three other guesses, numbered: "1: right to Ann. 2: write two Ann. Say choose and the number." Say "choose one", "choose two" or "choose three", and that guess takes the place of what was written. Only Windows speech recognition keeps other guesses; the other engines give one answer each, and "correct that" tells you so. Then say "scratch that" and say it again, or add a correction in My Words and Phrases.
Moving and selecting by voice
You can move the cursor and select text just by saying the words you are after. Say these at the start of a phrase, after a pause.
- "select" and the words. "Select the cat" selects the nearest place those words appear, looking back from the cursor first and then forward. Whatever you say next takes their place.
- "select" one set of words "through" another. "Select dear Sam through thank you" selects everything from the first words to the end of the second.
- "go to" or "go before" and the words puts the cursor just before them. "go after" puts it just after.
- "correct" and the words selects them, and you hear "Correcting the cat. Say the new words". Then say what should be there.
- "select again" or "select next" moves to the next place the same words appear, and "select previous" to the one before.
- Said on their own: "select sentence", "select line", "select paragraph", "go to start of paragraph" and "go to end of paragraph".
Capitals, punctuation and accents do not matter when dictation looks for your words, and numbers match either way, so "two" finds 2. Your corrections in My Words and Phrases are applied first. You hear once where you landed: "Selected: the cat", "Before the cat" or "After the cat". If that is three lines or more from where you were, the line number is added: "Selected: the cat, line 40".
If the words are not in the document, nothing moves. Your phrase is written as ordinary text, and you hear "Not found, written as text." Say "scratch that" to take it back out. All of this works with every speech engine, and the exact commands such as "select that" and "go to top" still do what they always did.
Letters, numbers, symbols and Markdown
Every mark you can say is in the tables above. A few ways of writing are worth a word of their own.
Capitals and spacing
- "caps on" starts every word with a capital until you say "caps off". You hear "Caps on." "Capitals on" and "capitals off" work too.
- "all caps on" writes every letter as a capital until "all caps off".
- "no space on" joins your words with no spaces and no automatic capitals, which is handy for web addresses and file names, until "no space off".
- Each lasts until you turn it off or stop dictation, and the status bar shows which one is on.
Spelling
Spelling mode now writes punctuation with no spaces, so "jay dot smith at sign example dot com" gives you an email address. To spell a single word without switching spelling mode on, say "spell" and the letters: "spell bravo alpha delta". If dictation got a word wrong, say "spell that": the phrase is selected, and the next thing you say is spelled out in its place. There is more in Spelling a word out.
Markdown and code
- "backtick" (or "back quote") opens or closes inline code, depending on how many are already on the line. "Code fence" (or "triple backtick") writes three backticks on a line of their own.
- "tilde", "vertical bar" (or "pipe symbol"), "caret", "greater than sign" and "less than sign" write those characters.
- At the start of a phrase, "bullet" or "list item" starts a bulleted line, "numbered item" a numbered one, "block quote" a quotation, and "heading one" to "heading six" a heading at that level. If the cursor is not at the start of a line, a new line is started first. In the middle of a sentence these stay words, so "the heading two lines down" is written just as you said it.
Snippets and clips by voice
Say these on their own, after a pause:
- "copy all" (or "copy everything" or "copy document") copies the whole document, just like Copy All.
- "copy that" copies the last phrase you dictated.
- "show clips" (or "show copy tray" or "open copy tray") opens the Copy Tray: the Paste from Tray list in QUILL Lite, and the Copy Tray window in QUILL.
- "show snippets" opens your snippets: Snippets in QUILL Lite, and Insert Snippet's list in QUILL.
And start a phrase with these:
- "paste clip" and a number from 1 to 12. "Paste clip three" writes what is in slot 3 of the Copy Tray; "paste slot three" works too. If the slot is empty you are told so, and nothing is written.
- "insert snippet" and its name. You do not need the exact name: capitals, punctuation and spaces do not matter, and the start of the name or one word in it is enough when only one snippet matches. If several match, you hear how many and the list opens so you can choose. If none match, you hear "No snippet called" and the name, and nothing is written.
- "insert abbreviation" or "expand abbreviation" and the abbreviation writes what it stands for.
A snippet or a clip goes in as one phrase, so one Ctrl+Z or "scratch that" takes it out again. The read-back just names it, "Snippet sign off" or "Pasted slot 3", rather than reading the whole thing. In QUILL, a snippet with blanks to fill writes each blank as its name in square brackets, like [name], so you can say "select name" and then say what goes there. None of these work while "Just write what I say" is on, because every command is written as words then.
Dictating in Spanish
New in QUILL Lite 1.2 and QUILL 1.0: you can dictate in Spanish. It is new, so think of it as something to try, and tell us how it goes.
To switch it on, open Dictation Settings (Alt+Shift+F6), move to Dictation language, choose Spanish, and press Enter. Choose English the same way to go back. Quicker still, press Ctrl+Shift+F11 or say "switch to Spanish", as described in Switching between English and Spanish.
What works now
- Your words come out in Spanish, accents and all, using Whisper's multilingual model. It comes with the app, so there is nothing to download. Moonshine only knows English, so in Spanish you get Whisper either way.
- Windows speech recognition needs Spanish installed in Windows: add Spanish (Spain or Mexico) under Windows Settings, Time and language, Speech. If it is missing, dictation tells you when you start.
- The wake and stop phrases become "Quill dicta" and "deja de dictar". If you typed your own, yours stay.
- Commands stay in English for now. "Scratch that", "select that", "new paragraph" and the rest work as they do in English. Spanish commands come once a native speaker has checked them. The one exception is switching back: "cambiar a inglés" and "dictado en inglés" work already.
Punctuation in Spanish
With automatic punctuation on (the usual setting), Whisper adds the commas, full stops and question marks for you, so just talk. The Spanish punctuation words only work when automatic punctuation is off, or with Windows speech recognition, because "coma" and "punto" are everyday words too. When they are on, you can say:
- "coma" for a comma, "punto" (or "punto y seguido") for a full stop, and "punto y aparte" for a full stop and a new paragraph
- "punto y coma", "dos puntos" and "puntos suspensivos"
- "abrir interrogación" and "cerrar interrogación", and "abrir exclamación" and "cerrar exclamación"
- "abrir paréntesis", "cerrar paréntesis", "abrir comillas", "cerrar comillas"
- "guion", "guion largo" and "arroba"
- "nueva línea", "nuevo párrafo" and "tabulador"
Say "literal" first to write the word itself: "literal coma" writes coma. Say "what can I say" while dictating to hear the full list. If a word keeps coming out wrong, use Help > Get Help from Support and tell us what you said and what was written.
Switching between English and Spanish
Press Ctrl+Shift+F11 (Switch Dictation Language) to change between English and Spanish, even in the middle of dictating. Or just say it: "switch to Spanish" or "Spanish dictation" while you are dictating in English, and "cambiar a inglés" or "dictado en inglés" while you are dictating in Spanish. The accents are optional.
You hear "Español." or "English." and carry on talking. The language you switch to is saved, so next time dictation starts in the one you used last. If you use Moonshine, the first switch to Spanish takes a few seconds while Whisper gets ready; after that, switching back and forth is quick until you close the program.
Dictation never guesses the language for you. It writes in the one you chose until you switch, so a French word in an English sentence does not send it off in the wrong direction.
Teaching dictation your words
Press Alt+Shift+F10, or the My Words and Phrases button in Dictation Settings. A window lists everything dictation has been taught, one line each, and the next phrase you say uses whatever you change. Each change is saved the moment you make it and read back to you. There are three kinds.
- Words Names, jargon and acronyms, spelled the way you want them. When the engine writes something that sounds or looks close, it is corrected to your spelling: add Tucson and "tuxon" becomes Tucson. The match is by spelling and by sound, so split names, run-together words and "R and D" heard for R&D are all repaired, while ordinary words are left strictly alone.
- Phrases Something short you say that writes something longer. Add Phrase asks what you will say (my email address) and what it writes. Press Enter in the second box for a new line, so a signature can be two lines.
- Corrections What the engine keeps hearing wrong, and what to write instead: quill light becomes QUILL Lite. Applied to every phrase, whatever the capitals, on every engine.
Edit (or Enter on a line) changes the line you are on and
Remove takes it out. Open the File opens the same
list as a plain text file, dictation.md, if you would rather edit a
file. Your own phrases appear in the "what can I
say" list with everything else.
In QUILL, Locked Dictation and Dictate (Offline) use the same list, so your words are spelled your way there too.
Starting and stopping by voice
The wake phrase
Switch on Listen for the wake phrase while dictation is off in Dictation Settings and you can start dictation by saying it instead of pressing a key. It is "Quill dictate" unless you choose another.
- Say the wake phrase, pause, and start talking. Or say it and keep going: "Quill dictate, dear Sam, thank you for your letter" wakes dictation and writes Dear Sam, thank you for your letter.
- It only counts at the start of what you say, so "I told Quill to dictate this" in conversation starts nothing.
- A near miss still works: "Quil dictate" wakes it.
- Your own phrase needs at least two words. A single word is refused, because it would start dictation by accident.
What it means for the microphone, exactly. While the wake phrase is on, the microphone is open whenever the program is the window in front. Nothing it hears is written, kept or sent anywhere unless it begins with the wake phrase; what does not is dropped the moment it has been checked. When another program comes to the front the microphone closes, and dictation stops if it was running. It does not listen while the program is in the background, so it will never pull you away from what you are doing in another program. The wake phrase is off until you turn it on.
It works with Moonshine, Whisper and Windows speech recognition, but not with Windows voice typing, which Windows runs itself.
The stop phrase
Say "stop dictation" on its own, after a pause, and dictation stops. You can choose your own words as well, in Dictation Settings; "stop dictation" keeps working either way.
- It counts only when it is everything you said in that breath. "I will stop dictation for today" in the middle of a paragraph is written, not obeyed.
- Like the wake phrase it forgives a near miss and needs at least two words.
- With the wake phrase on, stopping goes back to listening for the wake phrase, so you can start and stop without touching the keyboard at all.
Longer pauses, and sentences that carry on
If you like time to think between phrases, open Dictation Settings and set Pause before a phrase is written to Longer (2 seconds) or Longest (3 seconds). Each phrase then waits that long after you stop talking before it is written. Your words appear a little later, but a thinking pause no longer cuts you off.
Dictation also notices when you have not finished a sentence. If a phrase ends on a word a sentence hardly ever ends on, such as "a", "the", "of", "to", "and", "my" or "very", the full stop the engine put there is taken back out and your next phrase carries on the same sentence. Say "this is a very", pause, then say "good test", and you get This is a very good test. A full stop you said yourself is never taken back.
Dictation Settings, every option
Alt+Shift+F6. The window has two columns: what is heard and written on the left, starting and stopping on the right.
| Option | What it does | Starts as |
|---|---|---|
| Speech engine | Moonshine, Whisper, Windows speech recognition, Windows voice typing, any model you downloaded, or OpenAI with your own key | Moonshine |
| Automatic punctuation | The engine puts in the marks you do not say. Moonshine and Whisper only; off, you say every mark | On |
| Language | Which installed Windows speech language to use. The other engines understand English | The Windows default |
| Microphone | Which microphone to listen on, kept by name so it survives a change of default | The Windows default |
| Test Microphone | A button. Four seconds of listening, then how loud you were and, with Moonshine or Whisper, exactly what it heard. Nothing is kept | Not a setting |
| After each phrase is written, give me | A sound, speech (the words read back), both, or neither | Both |
| Saying "dash" writes | An em dash with no spaces, a spaced en dash, or two hyphens | Em dash |
| Pause before a phrase is written | Short (half a second), Normal (under a second), Long (about a second and a half), Longer (2 seconds) or Longest (3 seconds). Choose a longer one if it cuts you off while you are still thinking; the words appear a little later | Normal |
| Remove filler words like um and uh | Hesitations are left out instead of written. Real words are never removed | Off |
| Just write what I say | A pause does nothing at all: no full stop, no tone, no read-back, and voice commands are written as words. Spoken punctuation and the stop phrase still work | Off |
| Play sounds when dictation starts, stops or fails | The start, stop and error tones | On |
| Say "Dictation on" and "Dictation off" | The words that go with those tones | On |
| Listen for the wake phrase while dictation is off | Start dictation by voice | Off |
| Wake phrase | The words that start it. At least two words | Quill dictate |
| Stop phrase | The words that stop it, said on their own. At least two words | stop dictation |
| Stop dictation after silence | Never, or after 1, 5 or 10 minutes of hearing nothing, so dictation is not left writing in an empty room | Never |
Three buttons sit below them: Dictation Commands opens the full spoken list, My Words and Phrases saves these settings and opens the window described above, and More Dictation Settings (Alt+A) opens the window below.
More Dictation Settings, every option
The choices nobody needs on the first day. What you change here is saved when you press OK in Dictation Settings.
| Option | What it does | Starts as |
|---|---|---|
| Hold the dictation key to talk; a quick press still turns it on and off | Also lets you hold Ctrl+F11 while you talk and let go to stop | Off |
| While you speak, the words heard so far | Show them in the status bar and on braille, also say new words quietly, or do not show them | Show them |
| Send a dictated message to the AI | When I pause, or When I press Enter | When I pause |
| Pause before the message is written | Short, Normal or Long, when talking to the AI | Long |
| Remove filler words when talking to the AI | Leave um and uh out of what you say to the AI | On |
| Automatic punctuation when talking to the AI | Let the engine punctuate what you say to the AI | On |
| Let OpenAI dictation send my speech to OpenAI | Your agreement, which you can take back here | Off |
| OpenAI speech model | The model OpenAI dictation uses, from OpenAI's own list for your key | The newest |
| My Dictation Instructions | A button. Opens the file Tidy Dictated Text follows | Not a setting |
| Say punctuation marks in the read-back | The read-back says each mark by name, "Hello comma world period", whatever your screen reader's punctuation level | Checked |
| Time stamps in live transcripts | Each paragraph of a live transcript starts with the time, like [10:42] | Unchecked |
| Dictate in Other Programs | A button. Saves these settings, hands them to Quill Inkwell and starts it, so you can dictate into any program | Not a setting |
Knowing what dictation is doing
- The status bar has a Dictation part: Off, Listening, Hearing you, Writing, Spelling, or Waiting for wake phrase. Press F6 to reach the status bar and move to it; Enter there starts or stops dictation.
- The menu row is checked while dictation is writing.
- The tones and the words say when it starts and stops.
Dictation Status: asking what it is doing
Press Alt+F9 at any time to hear what dictation is doing, without changing a thing. With dictation off you hear something like "Dictation is off. Moonshine, in English." With it on you hear the engine and the language, and whether caps or no space is on: "Dictation on, Whisper, in Spanish. Caps on." During a live transcript you hear how long it has been running and how many words it has.
In QUILL Lite it is Dictation Status in Tools, Dictation. In QUILL, Alt+F9 answers for Locked Dictation while that is recording, and for live dictation the rest of the time.
Seeing the words while you speak
With Nemotron (an optional download) or OpenAI dictation, the program knows your words while you are still saying them. The status bar shows them after the word "Hearing", and a braille display shows them as they arrive. They are not in your document yet, because the engine may still change its mind; when you pause, the final words go in and the preview goes away. Nothing is spoken, so nothing talks over you. To hear new words quietly as well, choose Show it, and say new words quietly under While you speak, the words heard so far in More Dictation Settings.
Words go where you started
Into the document you started dictation in, at the cursor, replacing any selection. Your words go where you were when you started speaking: if you move the cursor, or switch to another window, while a phrase is still being recognised, it is still written there, your cursor is put back where you moved it, and you hear "Written where you started". If that spot changed meanwhile, the words go at the cursor and you are told. After that, dictation stops by itself when that document closes, when you switch it between plain and rich text, or when you speak while another program or a dialog is in front, and the program tells you why it stopped.
There is one microphone, so pressing Ctrl+F11 in a second document while dictation runs in the first moves dictation there and says so: "Dictation moved to Letter.txt". Dictation will not start in a read-only document, such as the read-only copy Compare opens, and says so before the microphone is opened.
If the microphone drops
If the microphone is unplugged, a Bluetooth headset drops, or the input goes completely silent for a few seconds, dictation says "The microphone stopped. Dictation is paused and will resume when it comes back." It stays on, writes nothing, and watches for the device; when it returns you hear "Microphone back. Listening." and carry on. The status cell says "paused, microphone lost" meanwhile.
If the speech engine itself stops answering it is restarted once, silently, and you notice nothing. Only a second failure is reported, by name, with the engine to try instead: "Moonshine stopped working. Try Whisper in Dictation Settings."
Dictating into Find, Replace and the AI box
Ctrl+F11 in the Find box, either Replace box, the question box of the AI pad, or the message box of the AI Conversation window dictates into that box: the same engine, the same words you have taught it, and Escape still cancels a phrase. "New line" and "new paragraph" become a space in a one-line box, no closing full stop is added, and the commands that move around a document do nothing there. Press Ctrl+F11 again to stop.
Dictating into other programs with Quill Inkwell
Dictate Anywhere lets you dictate into your email, a web page, a chat window or any other program, with the same engines, the same words you have taught it and the same voice commands. It lives in Quill Inkwell, the free family app that already sits quietly in the background and types into other programs for you.
- Open Quill Inkwell and choose File, Dictate Anywhere Key. Press the key you want to use anywhere in Windows. There is none until you choose one, so it can never clash with a key you already use.
- Go to the program you want to write in, and put the cursor where the words should go.
- Press your Dictate Anywhere key, speak, and pause. Each phrase is typed where the cursor is.
- Press the key again, or say "stop dictation", to stop.
The easy way to start: in QUILL or QUILL Lite, open More Dictation Settings and press Dictate in Other Programs. Your dictation settings are handed to Quill Inkwell and it starts; if you have not chosen a key yet, the key chooser opens for you. After that, Inkwell keeps its own copy of the settings, which you can change in its Dictation menu under Dictation Settings. It uses the same My Words and Phrases list as the editors.
Punctuation, spelling, caps and no space, switching language and "scratch that" all work in other programs. "Scratch that" erases what was just typed, but only while you are still in the same window, so it can never erase somebody else's typing. Anything that needs to read the other program's text, such as selecting by voice, "go to", "correct", clips and snippets, cannot work there, and you hear "That works in QUILL's own documents. In another program, say the words again, or scratch that."
To keep things safe, nothing is typed into a password field ("That is a password field, so nothing is typed there."), into a program running as administrator when Inkwell is not, or into Inkwell's own window. In QUILL's own windows you are reminded to use Ctrl+F11, which already dictates there. The read-back starts as a tone only, because your screen reader already echoes what is typed. In Safe Mode, Dictate Anywhere is off.
Talking to the AI
You can have a spoken conversation with AI help. Open the AI Conversation window; the cursor is already in Your message.
- Press Ctrl+F11 and ask your question.
- Pause. Your message is sent, and you hear "Working."
- The reply is read aloud. While it is read, the microphone does not listen, so the reply is never taken as your next message.
- When it has been read, carry on talking. To speak sooner, press Escape: you hear "Listening."
- Press Ctrl+F11 again when you are done.
In that box dictation uses its Talking to AI settings by itself: a longer pause, so a breath does not send half a question, filler words removed, and automatic punctuation on. They are in More Dictation Settings, with Send a dictated message to the AI, which you can set to When I press Enter to check each message first.
OpenAI dictation, with your own key
Optional, and off until you choose it. Your computer's own engines stay the default.
With your own OpenAI key you can have OpenAI recognise your speech. It is very accurate, punctuates well, gets names right that it has never met, and can show your words as you speak. What you say is sent to OpenAI: each phrase goes over an encrypted connection with your own key, OpenAI bills your account, nothing goes through QUILL's servers, and QUILL keeps no copy of your voice. Silence is never sent. QUILL's free AI is never used for dictation, and it does not work in Safe Mode.
- Save your OpenAI key in Use My Own AI Key (Alt+F2), or from Dictation Settings: More Dictation Settings..., then Add or Change OpenAI Key... (Alt+K).
- Open Dictation Settings (Alt+Shift+F6) and choose OpenAI (your own key; sends your speech to OpenAI) as the speech engine.
- A question says exactly what is sent. Choose Yes to agree; No, the answer if you just press Enter, keeps your engine and sends nothing.
- More Dictation Settings opens on OpenAI speech model, listing the models OpenAI offers your key, newest first, with the newest chosen. Models OpenAI is retiring are left out. gpt-live-transcribe writes as you speak; gpt-transcribe sends each phrase when you pause.
- Press Enter twice to save, then dictate with Ctrl+F11 as before.
If OpenAI stops offering your model, you are told once and asked to choose another; dictation never changes it for you. The words in My Words and Phrases are sent as words to expect, so names come out right. To stop sending your speech, choose another engine, or turn off Let OpenAI dictation send my speech to OpenAI in More Dictation Settings.
Live transcripts
A live transcript writes down everything you say, for as long as you like, in a document of its own. It is made for a meeting, a lecture, or a long stretch of thinking out loud. Press Ctrl+Alt+Shift+PageDown (Start or Stop Live Transcript), or choose it from the dictation menu.
- A new untitled document opens, and you hear "Live transcript on, in a new document." It never writes into the document you were working on.
- Everything is written down as it is: no voice commands except the stop phrase, no tone or read-back after each phrase, filler words left out and punctuation added for you. It never stops by itself after a silence.
- A pause of four seconds or more starts a new paragraph. Check Time stamps in live transcripts in More Dictation Settings and each paragraph starts with the time, like [10:42].
- New words always go at the end of the transcript, even while you work in another document or another program, and your own place in the transcript is kept, so you can read back over it while it carries on.
- The status bar shows how it is going, as in "Live transcript: 12 minutes, 1,840 words", and Alt+F9 says the same.
To stop, press Ctrl+Alt+Shift+PageDown again, or Ctrl+F11. You hear "Live transcript stopped" and the number of words, such as "Live transcript stopped, 1,840 words." The document stays open, unsaved, for you to save where you like; until you do, it is protected like any other unsaved document. In QUILL you can then run a Transcript Action on it.
The first time you start one, you hear "Please record other people only when they have agreed." Live transcripts listen to your microphone only, not to sound playing on the computer, and they need one of the program's own engines rather than Windows voice typing.
Transcribing an audio file
A recording into text, in the background, with the same speech engines you dictate with. The same command, window and key in QUILL and QUILL Lite.
Sometimes the words you want are already in a recording: a voice note from your phone, an interview, a lecture, a meeting somebody recorded for you. Press Shift+F5 (in QUILL Lite, Tools, Dictation, Transcribe a Recording; in QUILL, Tools, Speech, Live Dictation, Transcribe a Recording), choose the file, and keep working while it listens.
- Type the path of the recording, or press Browse... and choose it. You can drop a file on the window too, and choose several in Browse to have them done one after another.
- Check the language: English or Spanish.
- Check the speech model. The most accurate one on your computer is already chosen, and the box under it says how long the recording is and about how long it should take.
- Choose where the text goes: a new document, which is already chosen, or this document, at the cursor, when it finishes.
- Press Enter. You hear when it starts, the percentage quietly at each quarter, and one sentence when it is done, such as "Transcribed meeting.mp3: 12 minutes, 1,804 words, in 6 minutes, in a new document."
Paragraphs start where the recording pauses for two seconds or more. Timestamps
are off unless you turn them on: check Add timestamps and each
paragraph begins with its time, like [00:01:23]. Spoken
"comma" or "new paragraph" are written as words, because in a recording people
usually mean the words; check Obey spoken punctuation and commands
for a recording you dictated on purpose. Your corrections from My Words and Phrases
are used, and filler words go if you remove them when you dictate.
Which model for which recording
- Voice notes and dictated letters, one clear voice: Moonshine tiny, built in, is the fastest and does this well.
- Interviews and meetings, several voices: a downloaded model is worth it. Parakeet TDT 0.6B v3 is our first choice; Whisper small, medium or large-v3-turbo are good too, just slower.
- Lectures and talks: Parakeet or a Whisper model, with the speaker's names added to My Words and Phrases first.
- Noisy rooms: the bigger the model, the better it copes. The two tiny models struggle with noise.
- Spanish: the built-in Whisper tiny, Parakeet TDT 0.6B v3, Nemotron and the multilingual Whisper models.
| Model | Modest computer (2 cores, 4 GB) | Capable computer (4 cores, 8 GB) |
|---|---|---|
| Moonshine tiny (built in) | 6 minutes | 3 minutes |
| Whisper tiny (built in) | 19 minutes | 9 minutes |
| Parakeet TDT 0.6B v3 | 30 minutes | 13 minutes |
| Nemotron | 37 minutes | 16 minutes |
| Whisper small | 1 hour 49 minutes | 46 minutes |
| Whisper large-v3-turbo | about 6 hours | about 2 and a half hours |
Files it reads: MP3, M4A, AAC, WAV, Ogg, Opus, FLAC and WMA, plus the sound in an MP4 and AIFF, in both products, with parts of Windows that are already on your computer. Nothing extra to download, and no ffmpeg needed. Stopping: press Shift+F5 again and choose Stop Transcribing. Every result, finished or failed, stays in Activity (Shift+F9).
Privacy. With the models on your computer, nothing leaves it. If your own OpenAI key is saved, OpenAI's speech models join the list, but they are never chosen for you, never offered in Safe Mode, and each recording is sent only after you say Yes to a question that says exactly what is sent. It goes about a minute at a time over an encrypted connection, OpenAI bills your account by the minute, and QUILL keeps no copy. Please only send recordings you have the right to share.
Tidy Dictated Text: tidying with AI, when you ask
Speech recognition writes what it heard, and what it heard is not always the word you meant: their for there, a name it has never met, two words run together, a comma where a full stop belonged, and every "um" you did not know you said. Ctrl+F3 fixes exactly that and nothing else.
- Dictate as usual. When the paragraph is done, leave the cursor in it, or select exactly the stretch you want tidied.
- Press Ctrl+F3. You hear "Working." The selection, or else the paragraph the cursor is in, goes to the model with one instruction: correct misheard words, punctuation and capitalisation, remove fillers and false starts, and change nothing else.
- The Tidied Dictation window opens with focus on the corrected text, so you hear it first. Replace My Selection puts it where the dictated text was, Copy puts it on the clipboard, and Escape keeps what you had. If you typed in that paragraph while the answer was on its way, Replace is not offered and the window says so.
- Ctrl+Z takes a replacement back, as one step.
Nothing changes until you press Replace, and the model is told to add nothing. It needs a ChatGPT subscription or your own AI key (OpenAI or Google Gemini). QUILL's free AI service does not offer it. If you have neither, it tells you and opens the account window.
My Dictation Instructions
Tell Tidy Dictated Text how you like your writing. In Dictation Settings press More Dictation Settings (Alt+A), then My Dictation Instructions (Alt+I). A short file opens; write one instruction per line, such as "Write numbers as digits" or "Use British spelling", and save it. They go with the passage you tidy, and only then. The AI is told your dictated text is something to correct, never a request to answer.
Dictation context for each document
Tell dictation what a document is about, and OpenAI dictation and Tidy Dictated Text can use that to get your words right. Press Ctrl+Alt+Shift+PageUp (Dictation Context for This Document). The window has four parts:
- This document is: describe it in your own words, such as "a letter to my landlord about the boiler" or "notes from the garden club".
- Start from a saved context: the contexts you have saved, then a few to start from: Formal letter, Note to a friend, Technical writing, Meeting notes and Story.
- Also save it as a context named: give it a name to use it again in other documents.
- Who uses it: a read-only box that tells you which parts of dictation will use it.
Each document remembers its own context. An untitled document keeps it until you close the program. While dictating, say "dictation context" and the name of a saved one, such as "dictation context meeting notes", to choose it by voice.
OpenAI dictation uses the context to understand you better, and Tidy Dictated Text uses it alongside My Dictation Instructions. Moonshine, Whisper and the Windows engines cannot use it, and the window tells you so. Your context only ever goes where your speech or your text is already going.
What leaves your computer
Dictation sends nothing unless you choose OpenAI dictation with your own key and agree to it. One optional command, Tidy Dictated Text, sends a passage when you press it.
| Part of dictation | Where it happens | What is kept |
|---|---|---|
| Recognising what you said | On this computer, by an engine inside the program, unless you chose OpenAI | Nothing. The words are recognised, written, and forgotten |
| OpenAI dictation, if you choose it | Sent to OpenAI, with your own key: only the speech dictation heard, never silence | QUILL keeps nothing. OpenAI's own policies apply to what it receives |
| The list of OpenAI models | Asked of OpenAI with your key when More Dictation Settings opens, on a computer with an OpenAI key saved | Nothing of your writing or speech is sent |
| The audio from your microphone | On this computer, in memory | No recording is keptQUILL's Locked Dictation saves a take to a recovery folder until its transcript is safely inserted, then removes it |
| The words you teach it | A plain text file in your own data folder | Your dictation.md, until you change it |
| Recent phrases | In memory only | The last twenty of this session, gone when the program closes |
| Dictation context for a document | A small file beside your dictation.md. Sent only with OpenAI dictation or Tidy Dictated Text, along with what they already send | Your contexts, until you change them. An untitled document's context is forgotten when the program closes |
| Live transcripts | On this computer, written into a new document, unless you chose OpenAI dictation | The document, unsaved until you save it |
| Dictate Anywhere, in Quill Inkwell | On this computer, typed into the program in front, unless you chose OpenAI dictation | Nothing. Inkwell keeps only its copy of your dictation settings |
| Listening for the wake phrase | On this computer, and only while the program is in front | Nothing, unless what you said begins with the wake phrase |
| Test Microphone | On this computer | Nothing |
| Transcribing a recording | On this computer, unless you choose an OpenAI model and say Yes to the question for that recording | Only the text you get back. The recording is read, never copied |
| Tidy Dictated Text | Sent to OpenAI or Google, through your own subscription or key, with My Dictation Instructions if you wrote any | Whatever your own agreement with them says. Nothing goes through QUILL's servers |
Nothing above happens in the background or as you type. Dictation opens the microphone when you ask it to and closes it when you stop. The wake phrase is the one setting that keeps the microphone open. It is off until you turn it on, it closes the microphone whenever another program comes to the front, and it never listens while the program is behind another window.
On a modest computer
Nothing runs while dictation is off: the speech model loads the first time you start, stays ready while you use it, and is put away a few minutes after you stop, giving its memory back. If an optional model you downloaded cannot keep up with your voice, because the computer is busy or on battery saver, you hear "This computer is busy, so dictation switched to the faster built-in engine for now. Your choice in Dictation Settings is unchanged." Next time, it tries your model again.
Dictating well
A few habits make a big difference, whichever engine you use.
- Speak in whole phrases. Say a sentence or a clause the way you would to a person, then pause. Engines understand a whole phrase far better than one word at a time.
- Pause before and after a command. "Scratch that", "select the cat" and the rest only count as a phrase on their own, so a short pause either side is what tells dictation you mean it.
- Need time to think? Choose a longer pause. Longer or Longest in Dictation Settings, so a thinking pause does not end your sentence. If it does, and the sentence ended on a word like "the" or "and", dictation joins it back up for you.
- Fix words by saying them. "Correct" and the wrong words, then the right ones, is often quicker than reaching for the keyboard. "Spell that" is there for a name the engine has never heard.
- Teach it once. A name it keeps getting wrong belongs in My Words and Phrases (Alt+Shift+F10), so you never correct it again.
- Use a headset. A headset microphone beats a laptop's built-in one, and keeps the read-back out of the microphone.
- Lost track? Press Alt+F9 to hear what dictation is doing, or Shift+F11 for the last twenty phrases.
If something goes wrong
Dictation always says what happened and what to do about it. These are the usual reasons, and the fix for each.
It will not start
- No microphone was found. Connect one and press Ctrl+F11 again.
- The microphone you chose is not connected. Plug it in, or choose another in Dictation Settings. It will not quietly listen on a different one.
- The microphone could not be opened. In Windows Settings, go to Privacy and security, then Microphone, and make sure desktop apps are allowed to use it.
- A speech engine is not included in this copy. Choose another engine in Dictation Settings; reinstalling puts the missing one back.
- No speech recogniser is installed, with Windows speech recognition. Add a speech language in Windows Settings, Time and language, Speech.
- "This document is read-only, so dictation cannot write here." Dictate into a document you can type in.
It keeps mishearing me
- Try Whisper instead of Moonshine, or the other way round. They fail on different voices.
- Use a headset microphone rather than a laptop's built-in one. This is the single biggest improvement available to most people.
- If you are using speakers, set the read-back to a sound only. Otherwise the microphone hears the read-back and writes it down again.
- Add the names you use often to My Words and Phrases (Alt+Shift+F10), and add a correction for anything it gets wrong the same way every time.
- Choose a Long pause if it cuts you off mid-thought.
- Tidy the result afterwards with Ctrl+F3, if you have a ChatGPT subscription or your own OpenAI or Google Gemini key.
It stopped on its own
- "Moonshine stopped working. Try Whisper in Dictation Settings." The engine failed twice in one session; the first time it was restarted for you without a word. Choose the other engine, and if it keeps happening, write to support.
- After a few minutes of silence. That is "Stop dictation after silence" doing its job; it says so, with the number of minutes. Set it to Never in Dictation Settings if you would rather it waited.
- When you moved to another program. Dictation follows the document, not the screen. It tells you why it stopped.
A command was written as words instead of obeyed
A command counts only as a whole phrase, on its own, after a pause. Said inside a sentence it is just words. That way you can still dictate the sentence "delete that line of text". Check also whether "Just write what I say" is switched on, which writes every command as words by design.
If you hear "Not found, written as text." after "select" or "go to", the words you asked for are not in the document, so your phrase was written instead. Say "scratch that", then try again with words that are there.
I said "scratch that" one time too many
Press Shift+F11 for the last twenty phrases and insert the one you lost. That is what the list is for.
Which product has which dictation
Live Dictation is identical in both, down to the keys. Everything QUILL adds is a second, different way to dictate.
| Capability | QUILL | QUILL Lite |
|---|---|---|
| Live Dictation | YesCtrl+F11, same code, same keys | YesCtrl+F11, same code, same keys |
| The four Live engines | Yes | YesMoonshine and Whisper ship inside the installer |
| Spoken commands and spelling mode | Yes | Yes |
| Wake phrase and stop phrase | Yes | Yes |
| Words, phrases and corrections | Yesthe same file also feeds Locked Dictation | Yes |
| Recent phrases | Yes | Yes |
| Tidy Dictated Text and My Dictation Instructions | YesCtrl+F3 | YesCtrl+F3 |
| Holding Ctrl+F11 to talk | Yes | Yes |
| Seeing the words while you speak | Yeswith Nemotron or OpenAI | Yeswith Nemotron or OpenAI |
| Correct that | Yes | Yes |
| Talking to the AI | Yes | Yes |
| OpenAI dictation, own key | Yesoff until you choose it | Yesoff until you choose it |
| Selecting, moving and correcting by voice | Yes | Yes |
| Snippets, clips and abbreviations by voice | Yes | Yes |
| Switching between English and Spanish | YesCtrl+Shift+F11 | YesCtrl+Shift+F11 |
| Live transcripts | YesCtrl+Alt+Shift+PageDown | YesCtrl+Alt+Shift+PageDown |
| Dictation context for each document | YesCtrl+Alt+Shift+PageUp | YesCtrl+Alt+Shift+PageUp |
| Dictation Status | YesAlt+F9 | YesAlt+F9 |
| Dictate in Other Programs, through Quill Inkwell | Yes | Yes |
| Optional speech models (Speech Models) | Yes | Yesnothing downloads unless you ask |
| Locked Dictation | YesCtrl+F9, records then transcribes, with a review list | No |
| Dictate (Offline) toggle | Yes | No |
| Parakeet 3 for Locked Dictation | YesManage Speech Models | NoLocked Dictation is QUILL's own |
| Transcribing audio and video files | Yesincluding captions, speaker labels, and a watch folder that transcribes what you drop in | No |
| Voice commands and conversation mode | Yesa safe allowlist; dictation writes, voice commands act | No |
A dashed line down the left of a row means one of the two does not have it. For everything else the two products differ on, see how QUILL and QUILL Lite compare, capability by capability.
Downloads and further reading
Dictation is in both programs at no cost and with nothing to add. In QUILL Lite it is on in every feature profile except WordPad and Notepad, and can be switched off in Customize Features like any other area. Switching it off also removes its menu, its keys and its wake phrase.
Get either program
QUILL Lite 1.1.2 installer for Windows All QUILL 0.9.0 Beta 2 downloads, including macOS and portable
Windows 10 or later, 64-bit for dictation in either program. Dictation only works on Windows, so QUILL on a Mac has everything else but not dictation. QUILL Lite's installer is about 230 MB, mostly because the two speech engines are inside it.
Read more
- The QUILL Lite user guide: dictation alongside everything else the program does.
- The QUILL user guide: Locked Dictation, transcription, captions and voice commands.
- How QUILL and QUILL Lite compare: everything that is different between the two.
- The QUILL Lite AI guide: what Tidy Dictated Text sends, and what the free service does.
- QUILL frequently asked questions: installing, updating and privacy.