Transcription
- Whisper large-v3-turbo by default, on Metal with CoreML
- Model sizes from 80 MB to 1.55 GB, including a 600 MB quantised build
- Parakeet TDT 0.6B on CPU for machines without a usable GPU path
- 16 kHz mono resampling handled before the model sees anything