Back to Add-ons
Local Models (sherpa-onnx)
Bundledby TypeWhisper
TranscriptionWindows
About
Local Models (sherpa-onnx) adds offline transcription to TypeWhisper for Windows using sherpa-onnx. The plugin currently exposes NVIDIA Parakeet for broad local transcription and Canary 180M Flash for a smaller model with translation support.
Features
- Offline transcription on Windows through sherpa-onnx
- Downloadable model files from Hugging Face
- Automatic, CPU, or NVIDIA CUDA acceleration
- Managed CUDA runtime installation with a safety probe and CPU fallback
- Download, load, unload, and remove controls for local models
- Parakeet model with broad language coverage
- Canary model with English, German, French, and Spanish support
- Translation support when the Canary model is selected
- No API key or cloud account required after the model is downloaded
Transcription Models
| Model | ID | Size | Languages | Translation |
|---|---|---|---|---|
| Parakeet TDT 0.6B | parakeet-tdt-0.6b | ~670 MB | 25 | No |
| Canary 180M Flash | canary-180m-flash | ~198 MB | 4 | Yes |
Configuration
- Model - Select either Parakeet TDT 0.6B or Canary 180M Flash.
- Download - Downloads all required ONNX model files and token files.
- Language - Canary supports
en,de,fr, andes. - Acceleration - Use automatic selection, CPU, or NVIDIA CUDA. A changed runtime may require reloading the model.
Setup
- Open TypeWhisper Settings > Integrations
- Select Local Models (sherpa-onnx)
- Download and load the model you want to use
- Select the sherpa-onnx provider as your transcription engine
Notes
- Parakeet is the recommended model for general offline transcription.
- Canary is the option to use when you need translation support.
- If the CUDA runtime cannot be loaded safely, TypeWhisper reports the reason and falls back to CPU.
- The plugin stores model files under the plugin data directory in
Models/<model-id>/.