Text-to-Speech

Open-source text-to-speech models and voice synthesis engines

Text-to-Speech — comparison of VoxCPM, fish-speech, sherpa-onnx
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
SOTA Open Source TTS
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
Popularity
Stars38,50132,98815,201
Global Rank#799#1058#3274
Weekly Activity(Oct 3 – Oct 9)
New Stars+243+74+113
Pushes000
Issues Closed000
Community
Forks4,3492,8611,761
Contributors31104247
Open Issues12615665
Project Info
OwnerOpenBMBfishaudiok2-fsa
LicenseApache-2.0NOASSERTIONApache-2.0
LanguagePythonPythonC++
CreatedSep 2025Oct 2023Sep 2022