Replies: 2 comments 2 replies
|
This is a pretty big model and not really suited to most people without GPU's or lower end GPU's |
|
Based on my testing, Parakeet performs well for casual English dictation but loses accuracy on technical vocabulary, medical terminology, and some non-English languages. Whisper Large V3 Turbo with Vulkan runs well on my AMD Ryzen 7 7840HS system, although Handy's current Whisper integration is not yet reliable there. Canary may be another useful option for GPUs with approximately 12 GB of VRAM. VibeVoice appears more resource-intensive and may be better suited to long-form workloads. My current preference would be to stabilize Whisper support and benchmark Canary, Parakeet, and VibeVoice against the same representative workload before choosing a default direction. |
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
VibeVoice-ASR may be worth evaluating as an alternative local backend, particularly for long-form or multilingual transcription:
A useful comparison would measure latency, memory usage, technical-vocabulary accuracy, multilingual accuracy, and GPU requirements against Whisper, Parakeet, and Canary. VibeVoice may be too resource-intensive for Handy's typical hardware, but its architecture and results are relevant to future backend support.
All reactions