Back to Radar简体中文

jamiepine/voicebox vs modelscope/FunASR

Compare jamiepine/voicebox and modelscope/FunASR using the current verified snapshot: positioning, license, deployment, use cases, limitations, and original sources.

jamiepine/voicebox

A local-first voice I/O stack that combines dictation, text-to-speech, voice cloning, and agent voice output. It is built with Tauri (Rust) and routes audio processing through local models such as Whisper and Kokoro.

License
MIT
Deployment
Refer to project documentation
Use cases
Meeting Notes · Audio & Speech · Automation
Updated
2026-07-16T18:10:45Z

Original project link

modelscope/FunASR

FunASR is an industrial-grade speech recognition toolkit supporting offline, streaming, and edge deployment for ASR, VAD, punctuation, speaker diarization, and emotion recognition. It provides a multi-model alternative to Whisper and cloud APIs, offering up to 340x realtime transcription speeds across 52 languages while running on CPU and edge devices.

License
MIT
Deployment
Refer to project documentation
Use cases
Meeting Notes · Translation & Subtitles · Audio & Speech · Automation
Updated
2026-07-17T12:02:14Z

Original project link

How to choose

First eliminate options that fail required deployment, license, or use-case constraints; then inspect each detail page for limitations and direct evidence.