{"repo":"diodiogod/TTS-Audio-Suite","free":true,"listed":false,"github":"https://github.com/diodiogod/TTS-Audio-Suite","clone":"git clone https://github.com/diodiogod/TTS-Audio-Suite.git","description":"A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools","language":"Python","stars":1166,"topics":["ai-audio","audio-processing","chatterbox","comfyui","f5-tts","higgs-audio","text-to-speech","tts","voice-cloning","voice-conversion"],"license":null,"category":"media-processing","readme_excerpt":"[![Stargazers][stars-shield]][stars-url] [![Issues][issues-shield]][issues-url] [![Forks][forks-shield]][forks-url] [![Dynamic TOML Badge][version-shield]][version-url] TTS Audio Suite v5.8.3 Universal multi-engine TTS extension for ComfyUI - evolved from the original ChatterBox Voice project. A comprehensive ComfyUI extension providing unified Text-to-Speech, Voice Conversion, Audio Editing, and integrated RVC model training through multiple engines including ChatterboxTTS, DramaBox, F5-TTS, Higgs Audio 2, Higgs Audio v3, Step Audio EditX, MOSS-TTS, Echo-TTS, and RVC (Real-time Voice Conversion), with modular architecture designed for extensibility, runtime isolation for fragile legacy stacks, and a modern Transformers 5 main environment. Subtitle workflows are still a core focus: the suite can transcribe to SRT, rebuild subtitles from edited transcripts, or estimate fresh SRT timing from plain text using the same advanced readability rules, while preserving project control tags for downstream TTS. Quick Engine Comparison — 19 Engines Engine Languages Size Key Features -------- ----------- ------ -------------- F5-TTS 🇺🇸​🇩🇪​🇪🇸​🇫🇷​🇮🇹​🇯🇵 +4 1.2GB each Targeted Word/Speech Editing, Speed control ChatterBox 🇺🇸​🇩🇪​🇫🇷​🇮🇹​🇯🇵​🇰🇷 +4 4.3GB Expressiveness slider ChatterBox 23L 🌐 24 languages 4.3GB V1, V2, and V3 official checkpoints VibeVoice 🇺🇸​🇨🇳​🇩🇪​🇪🇸​🇫🇷​🇮🇹 +21 5.4GB / 18GB 90-min long-form, Native 4-speaker (Base models) Higgs Audio 2 🇺🇸​🇨🇳​","default_branch":null,"files":null,"tree":[],"storefront":"/r/diodiogod","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/diodiogod/TTS-Audio-Suite/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}