#59 · Primary category: Speech & Audio
EmotiVoice
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
Project last updated:08/13/24
GitHub Stars
8.5K
Forks
756
Contributors
13
License
Apache-2.0
Why we included this project
EmotiVoice generates spoken audio from plain text in English and Chinese, and its main selling point is emotional control: a prompt can push the output toward happy, excited, sad, or angry delivery instead of the flat reading most TTS engines produce. That makes it a good fit for narration, game dialogue, and any character-driven audio work. It also comes with a very large voice library, so you can match a character or brand tone without training your own model. There is a web interface for quick experiments, a scripting interface for batch runs, and an OpenAI-compatible TTS endpoint that slots into existing pipelines with little friction. Teams building spoken-audio products, interactive media, or localized voice apps should find it a practical place to start.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
whisper.cpp
Port of OpenAI's Whisper model in C/C++
Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
VibeVoice
Open-Source Frontier Voice AI
voicebox
The open-source AI voice studio. Clone, dictate, create.
TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production