#94 · Primary category: Speech & Audio

TTS-Audio-Suite

ai-audio audio audio-editing audio-generation audio-processing chatterbox comfyui cozy-voice-3 echo-tts f5 f5-tts higgs-audio indextts-2 qwen3-tts rvc text-to-speech tts vibevoice voice-cloning voice-conversion

A ComfyUI custom node suite for multi-engine, multi-language text-to-speech and voice conversion with SRT timing and audio tools.

Project last updated:08/28/26

GitHub Stars

1.2K

Forks

140

Contributors

13

License

Other

Why we included this project

ComfyUI users who want voice in their pipelines often end up installing a separate node pack for every model. TTS-Audio-Suite bundles roughly nineteen engines, RVC, Chatterbox, F5-TTS, Higgs Audio 2/3, IndexTTS-2, CosyVoice3, Qwen3-TTS, VibeVoice and others, behind one set of nodes with a shared interface, so a team that needs several languages or different synthesis styles avoids maintaining multiple integrations. The subtitle workflow is the part that stands out: it can transcribe audio to SRT, estimate fresh timing from plain text, rebuild subtitles after edits, and feed project control tags straight into the TTS nodes, which makes dubbing and long-form narration practical. Voice cloning and RVC training are included as well, letting you shape a custom voice and then generate with it in the same tool. Everything runs locally, so audio generation and model weights never leave your machine.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category