jamiepine/voicebox vs modelscope/FunASR
Compare jamiepine/voicebox and modelscope/FunASR using the current verified snapshot: positioning, license, deployment, use cases, limitations, and original sources.
jamiepine/voicebox
A local-first voice I/O stack that combines dictation, text-to-speech, voice cloning, and agent voice output. It is built with Tauri (Rust) and routes audio processing through local models such as Whisper and Kokoro.
- License
- MIT
- Deployment
- Refer to project documentation
- Use cases
- Meeting Notes · Audio & Speech · Automation
- Updated
- 2026-07-16T18:10:45Z
modelscope/FunASR
FunASR is an industrial-grade speech recognition toolkit supporting offline, streaming, and edge deployment for ASR, VAD, punctuation, speaker diarization, and emotion recognition. It provides a multi-model alternative to Whisper and cloud APIs, offering up to 340x realtime transcription speeds across 52 languages while running on CPU and edge devices.
- License
- MIT
- Deployment
- Refer to project documentation
- Use cases
- Meeting Notes · Translation & Subtitles · Audio & Speech · Automation
- Updated
- 2026-07-17T12:02:14Z
How to choose
First eliminate options that fail required deployment, license, or use-case constraints; then inspect each detail page for limitations and direct evidence.