The meaningful changes to how you speak, write, and keep your words close.
Recording shortcuts
Double-tap, let go, and keep dictating.
Two quick presses of Ctrl+Shift lock recording so you can let go of the keys. When a failed-insertion card is showing, a single tap re-inserts its text at your cursor.
Locked recording continues until another press, Escape, or the 12-minute limit.
The second short press must arrive within 350 ms. Without a failed card, a single tap still cancels a mis-trigger; a tap followed by a hold is normal dictation.
On Windows, switching keyboard layouts twice quickly with Ctrl+Shift can also trigger recording lock.
A first run that takes you all the way to dictation.
The redesigned setup walks through privacy, engine choice, microphone access, macOS Accessibility and a practice dictation. Leave midway and continue from the same step later.
Model download and permission setup can proceed alongside each other, with engine selection and readiness tracked separately.
New installations use Qwen3-ASR 0.6B by default except on machines detected as low-end, which default to OpenAI. Existing installations keep their selected engine.
OpenAI requires your own API key and sends audio to the provider. First-run setup no longer offers 1.7B up front; the larger model can be chosen later.
v1.15.5Windows microphone status now reflects actual recording attempts, with ready, denied, no-device and error states in setup, Home and Settings.
v1.17.0First run was simplified to local versus cloud without offering 1.7B up front. Groq and Nemotron moved under More while the active engine stays visible.
Local models
A larger local model, when you want the choice.
Qwen3-ASR 1.7B arrived as an experimental option in v1.15.0. Qwen is now one engine with two sizes, 0.6B and 1.7B, with separate downloads and switching suggestions based on your machine’s dictation speed.
Keep dictating with 0.6B while 1.7B downloads; SayType switches when it is ready unless you have chosen another engine in the meantime.
Returning to Qwen restores the size you last used. Once both sizes are downloaded, you can switch directly from Home.
1.7B is now described as the larger model option rather than experimental. Suggestions use measured dictation speed on your machine, not a promise of equal performance across devices.
v1.15.8Click a ready engine to activate it, or use its separate expand control to inspect settings without switching. Cloud model changes apply within the active engine.
v1.17.0Qwen became one engine with two sizes, gained switching suggestions based on measured speed, and dropped the experimental label for 1.7B.
v1.17.1Compact Home controls show Qwen’s size with a quick switch; computer and cloud icons distinguish local and cloud engines.
Recording
A native recording path on Mac.
Mac recording moved from the webview microphone API to CoreAudio to address quiet or garbled openings observed on affected Macs. The microphone is opened for each dictation rather than kept active between recordings.
The macOS microphone indicator is no longer kept on between dictations.
This is a capture change: it addresses how audio reaches the recognizer, rather than replacing the transcription model.
CoreAudio capture is macOS-only. Hardware behavior varies; this milestone does not claim every microphone or Bluetooth configuration has been verified.
v1.13.2Startup and shutdown waits are bounded, and interrupted recordings are identified explicitly so available text can be recovered.
v1.15.2Windows and Linux also open and close their webview microphone stream per dictation. Their recording backend remains different from macOS.
Local dictation
Less work left after a long dictation.
Qwen can process completed chunks while you keep talking, instead of waiting for the entire recording before starting transcription. That leaves less audio to process after you release the shortcut.
Finalized chunk text appears while the remaining audio is still being captured.
The engine starts warming up when recording begins. This is chunked processing, distinct from Nemotron’s live streaming recognition.
Applies to Qwen local dictation. Completion time depends on the recording and device; no fixed speed improvement is promised.
v1.9.1Recording finalization gained safeguards and recovery handling for missing or late recorder events.
v1.14.1Sample-coverage checks route detected gaps into recovery instead of automatically inserting a partial result. The original intermittent tail-loss trigger was still under investigation in this release.
Live transcription
See words appear while you speak.
Nemotron 3.5 added a second local engine with streaming recognition on Apple Silicon. It produces a live transcript during recording, alongside Qwen’s batch-oriented approach.
Each local engine has its own download and can be selected from the app’s engine controls.
Audio stays on-device when using the local engine. Translate mode was removed in v1.16.0.
Nemotron is an experimental option on Apple Silicon and Windows x64; platform availability is not a claim of end-to-end validation on every device.
v1.11.0Nemotron became available on Windows x64 using the CPU runtime.
v1.15.2The language picker became available for Nemotron rather than being locked to Qwen-style automatic detection.
v1.15.3Language changes now take effect in live dictation without restarting the app.
v1.16.0The separate translate mode, provider picker and upload-consent prompt were removed; local engines now hide the unneeded API key field.
Recovery
Retry a saved recording instead of saying it again.
Recovery first arrived for stalled local transcription: SayType retries once and, if it still cannot complete, saves the recording as a pending History entry for another attempt.
Later releases expanded saved-audio recovery to more transcription failures and improved the reasons shown in History.
A manual retry uses the engine currently selected, so changing provider or correcting an API key can make the saved recording usable.
Retry requires successfully saved audio. History has a 100-entry limit; saved recordings are not a permanent backup.
v1.13.3More failed requests retain audio and an error reason; retry uses the current engine and successful retries replace pending entries.
v1.13.4Empty retry results preserve the recording, and cloud or translation failures with incomplete capture no longer miss their History entry.
App updates
New versions, without another installer hunt.
Background update checks and in-app installation arrived with v1.4.0. SayType checks daily, downloads an available update, and lets you choose when to restart.
You can also check for updates from Settings.
Installing an update remains an explicit action rather than an automatic restart during your work.
This describes the built-in updater in official desktop releases.
v1.15.6The Home update card gained a What’s new link to that version’s GitHub release notes, so you can review changes before restarting.
Local transcription
Dictation that can stay on your device.
Qwen3-ASR brought on-device speech recognition to SayType. After downloading the model, local transcription can run without an account, API key or network connection.
Model download, progress and removal became available in Settings, alongside a choice between local and cloud transcription.
The next release made local setup more prominent on Apple Silicon and added quicker engine switching from Home and the menu bar.
Introduced for Mac, with Apple Silicon recommended for local use. Optional Groq/OpenAI transcription uploads audio to the selected provider; translate mode has been removed.