#59 · Primary category: Speech & Audio

EmotiVoice

ai deep-learning emotion emotivoice multi-speaker prompt python pytorch speech speech-synthesis style text-to-speech tts

EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine

Project last updated:08/13/24

GitHub Stars

8.5K

Forks

756

Contributors

13

License

Apache-2.0

Why we included this project

EmotiVoice generates spoken audio from plain text in English and Chinese, and its main selling point is emotional control: a prompt can push the output toward happy, excited, sad, or angry delivery instead of the flat reading most TTS engines produce. That makes it a good fit for narration, game dialogue, and any character-driven audio work. It also comes with a very large voice library, so you can match a character or brand tone without training your own model. There is a web interface for quick experiments, a scripting interface for batch runs, and an OpenAI-compatible TTS endpoint that slots into existing pipelines with little friction. Teams building spoken-audio products, interactive media, or localized voice apps should find it a practical place to start.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category