#63 · Primary category: Speech & Audio

TTS-WebUI

ace-step ai audio-generation cosyvoice generative-ai generator gradio music musicgen openai-api openvoice rvc styletts2 text-to-speech tortoise-tts tts vocos

A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Audio, MMS, StyleTTS2, MAGNet, AudioGen, MusicGen, Tortoise, RVC, Vocos, Demucs, SeamlessM4T, and Bark!

Project last updated:07/27/26

GitHub Stars

3.2K

Forks

329

Contributors

15

License

MIT

Why we included this project

TTS-WebUI puts more than twenty text-to-speech models behind one Gradio and React interface, so you install a single app instead of setting up Bark, GPT-SoVITS, CosyVoice, StyleTTS2, Piper, Kokoro, and XTTSv2 as separate tools. The same window also covers voice cloning and voice conversion through RVC, music and audio generation with MusicGen, MAGNeT, and Stable Audio, and cleanup work like Demucs. That reach makes it a practical base for a pipeline handling dubbing, character voices, narration, or game and chatbot audio. Since the model extensions are opt-in, a fresh install stays small and you load only what you need, keeping setup and memory overhead manageable.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category