On-device transcription should be the default, using local models like Whisper or Parakeet, so the feature remains fast, private, and free to run. This would reduce dependence on cloud processing, improve responsiveness, and give users a more secure experience with their voice data staying on the device.
