Text-to-Speech

Open-source text-to-speech models and voice synthesis engines

Text-to-Speech — comparison of VoxCPM, fish-speech, sherpa-onnx
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
SOTA Open Source TTS
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
Popularity
Stars35,94232,30914,298
Global Rank#856#1024#3458
Weekly Activity(Aug 14 – Aug 20)
New Stars0+2+1
Pushes000
Issues Closed000
Community
Forks4,1082,7841,643
Contributors3197230
Open Issues11320618
Project Info
OwnerOpenBMBfishaudiok2-fsa
LicenseApache-2.0NOASSERTIONApache-2.0
LanguagePythonPythonC++
CreatedSep 2025Oct 2023Sep 2022