#30 · Primary category: Speech & Audio
Bert-VITS2
vits2 backbone with multilingual-bert
Project last updated:08/24/26
GitHub Stars
8.8K
Forks
1.3K
Contributors
36
License
AGPL-3.0
Why we included this project
Bert-VITS2 gets its natural, expressive voice quality from pairing the VITS2 text-to-speech backbone with BERT embeddings. It trains on multiple languages, and the webui-based preprocessing and training workflow gives you fine control over voice characteristics, which makes it a good fit for teams that want to build custom voices, clone a speaking style, or drop speech synthesis into a larger product instead of paying for a closed commercial API. If you are comfortable fine-tuning PyTorch models, the code is approachable and clearly inherits from the open-source VITS and MassTTS lineage. One caveat worth knowing: the maintainers now point new users to their successor project Fish-Speech for ongoing development, so look there first if you are starting fresh.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
whisper.cpp
Port of OpenAI's Whisper model in C/C++
Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
VibeVoice
Open-Source Frontier Voice AI
voicebox
The open-source AI voice studio. Clone, dictate, create.
TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production