#94 · Primary category: Speech & Audio
TTS-Audio-Suite
A ComfyUI custom node suite for multi-engine, multi-language text-to-speech and voice conversion with SRT timing and audio tools.
Project last updated:08/28/26
GitHub Stars
1.2K
Forks
140
Contributors
13
License
Other
Why we included this project
ComfyUI users who want voice in their pipelines often end up installing a separate node pack for every model. TTS-Audio-Suite bundles roughly nineteen engines, RVC, Chatterbox, F5-TTS, Higgs Audio 2/3, IndexTTS-2, CosyVoice3, Qwen3-TTS, VibeVoice and others, behind one set of nodes with a shared interface, so a team that needs several languages or different synthesis styles avoids maintaining multiple integrations. The subtitle workflow is the part that stands out: it can transcribe audio to SRT, estimate fresh timing from plain text, rebuild subtitles after edits, and feed project control tags straight into the TTS nodes, which makes dubbing and long-form narration practical. Voice cloning and RVC training are included as well, letting you shape a custom voice and then generate with it in the same tool. Everything runs locally, so audio generation and model weights never leave your machine.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
whisper.cpp
Port of OpenAI's Whisper model in C/C++
Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
VibeVoice
Open-Source Frontier Voice AI
voicebox
The open-source AI voice studio. Clone, dictate, create.
TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production