Offline recognition lets you dictate tasks without an internet connection. On Android, it also helps when Google services are unavailable or system voice input can’t be used.
To use offline recognition, you only need to download a local speech recognition model once. After that, speech is processed directly on your device, and the audio isn’t sent to a server.
The task parameters that can be recognized—date, project, tags, priority, and reminder—stay the same. Offline recognition only changes how speech is converted into text, not how task parameters are extracted.
How to download and select a model
- Open Settings → Voice input.
- Tap Voice input method.
- On the next screen, download the model you need and select it. You can also keep using system voice input if it’s available.
There is a ? button to the right of the screen title. Tap it to see a short explanation of what local models are for and how to use them.
Downloaded models
The first section shows the voice input methods that are already available:
- System voice input — speech recognition provided by your device. On Android, this is usually Google; on iOS, it uses built-in system features. This option appears only if system speech recognition is available on the device.
- Local models — one row for each downloaded language. Below the language name, you’ll see the file size and the model’s technical name.
The selected method is marked with a green checkmark and the Selected label. To switch to another method, simply tap its row.
Downloaded models have a ⋮ menu on the right:
- if the model isn’t selected — Select and Delete;
- if the model is selected — Delete only;
- the selected system voice input method doesn’t have this menu.
If you delete the currently selected local model, the app will switch back to system voice input if it’s available.
Available device storage
The second section shows how much storage is free on your device and how much space is already used by downloaded models.
Before downloading another model, check this indicator—models can take up tens of megabytes.
Available for download
The third section lists languages whose models haven’t been downloaded yet.
Currently available:
Russian, English, Français, Deutsch, Português, Türkçe, Italiano, українська, Polski, қазақша, Español.
Tap a language row or the download icon on the right to start downloading. While the model is downloading, the icon is replaced with a progress indicator. You can cancel the download at any time.
If something goes wrong, an error icon will appear. Tap it to see the reason and try again.
Before downloading over a mobile network, the app will warn you that the download may use mobile data.
If there isn’t enough storage space, you’ll see Not enough space with a Retry button. If there is no network connection, you’ll see No network. If the downloaded file is corrupted, you’ll see Model file is corrupted and will need to download the model again.
While the model is downloading, progress is also shown in notifications. Once the model is ready, tap the notification to open the model selection screen.
If Google services aren’t available on your device
On some Android devices, such as devices without Google services, system voice input isn’t available. Previously, this meant that voice task creation couldn’t be used. Now you can simply download a local model.
The first time you try to dictate a task without a downloaded model, you’ll see the Set up voice input! window. It explains that you need to download a local model.
Tap Go to settings to open the voice input settings. The Voice input method option will be highlighted so you can go straight to model download.
If system speech recognition is unavailable and no local model has been downloaded, you may also see the Offline recognition unavailable screen inside the voice input window.
Tap Download offline model to open the settings. On Pro and Elite plans, this will usually take you directly to the model selection screen.
The screen closes automatically after 5 seconds or when you tap the back arrow.
Important! A local model is only used to convert speech into text. The voice input limit on the Free plan—10 recognitions—and unlimited voice input on Pro and Elite work the same way as with system speech recognition.