Skip to content

Audio Transcription

Configures automatic speech-to-text transcription for audio notes recorded in the Notepad Quick Entry. When enabled, audio notes are transcribed in the background on the server immediately after they are saved, and the note text is filled in with the result.

Everything runs locally — no internet connection or cloud service is involved.


The status card at the top of the page shows:

  • Notes waiting — audio notes that have not yet been transcribed and are waiting for a model to become available.
  • Notes queued — notes that were queued while transcription was unavailable. Use the Transcribe queued notes button to process them once a model is ready.
  • Active model — the model currently loaded by the transcription engine.
  • Engine status chip — whether the engine is running, idle, or in an error state.

If the engine encountered an error, the details are shown here. Errors on individual notes are shown as warnings and do not stop the engine.


Turns automatic transcription on or off. When enabled, every new audio note is transcribed as soon as it is saved. Notes recorded while transcription is disabled are not retroactively transcribed.

Select the speech recognition model to use. Two models are available:

ModelLanguagesSize
NVIDIA Parakeet TDT 0.6B v325 European languages with automatic language detection~490 MB
NVIDIA Parakeet TDT 0.6B v2English only — slightly more accurate for English notes than v3~485 MB

Models must be downloaded before they can be used. A chip next to each model indicates whether it has been downloaded.

Click Save & Restart Transcription to apply changes. This restarts the transcription engine with the new settings.


When a model is selected, the Model Management card shows its download status and size. If the model is not yet downloaded, a Download button initiates the download directly from the internet to the server. Progress is shown with a progress bar.

If a model is already downloaded, a Delete button removes it from disk to free space. Deleting the active model disables transcription until a new model is selected and downloaded.