#105 · Primary category: Speech & Audio

IMS-Toucan

deep-learning pytorch speech speech-processing speech-synthesis text-to-speech toolkit tts

Controllable and fast Text-to-Speech for over 7000 languages!

Project last updated:01/25/26

GitHub Stars

2.2K

Forks

316

Contributors

12

License

Apache-2.0

Why we included this project

IMS Toucan comes from the University of Stuttgart's Institute for Natural Language Processing and is the home of ToucanTTS, a text-to-speech system that covers more than seven thousand languages. It is fast and controllable, and inference does not need a GPU, so it works well for teams that want multilingual speech synthesis on modest hardware. The repository also doubles as a teaching resource, with code and recipes for training, fine-tuning, and understanding modern speech synthesis. A free GPU-backed demo on Hugging Face lets you hear the output before you commit, and the team published a companion multilingual dataset for anyone who wants to train their own models. The language coverage, the low compute needs, and the tutorial material together make it a strong pick for shipping a multilingual voice product or for studying how these systems work.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category