diodiogod/TTS-Audio-Suite
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools observed · 2026-09-01
Health v2 · maintenance only
84/100
- Activity 100
- Release rhythm 95
- Longevity 28
Flags: no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 7
- age_days: 392
- days_rel: 39
- days_push: 2
- n_releases_24m: 22
Adoption not part of the score
1185 stars · 141 forks observed · 2026-09-01
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A ComfyUI custom node suite providing unified multi-engine Text-to-Speech, Voice Conversion, and audio editing across 19 engines like ChatterBox, F5-TTS, Higgs Audio, and RVC. It also supports SRT subtitle workflows, character-based dialogue, and integrated RVC model training.
Use cases
- generate speech from text locally with multiple TTS engines
- clone a voice from a reference audio sample
- convert one voice to another with RVC
- generate character dialogue for videos with SRT timing
- transcribe audio to subtitles and rebuild them from edited transcripts
- edit speech emotion and style in existing audio
- build TTS workflows inside ComfyUI
When to choose
- you already use ComfyUI and want TTS or voice conversion nodes
- you need to compare or combine multiple TTS engines in one tool
- you need subtitle-synced, multi-character voiceover generation
- you want local, offline speech generation with voice cloning
When to avoid
- you need a simple standalone TTS CLI without ComfyUI
- you lack a GPU or the disk space for multi-GB models
- you need production cloud TTS with SLAs
- you want a lightweight plugin with minimal dependencies
Facets
plugin · maturity active
tts audio-processing speech-recognition machine-learning plugin-system speech-processing artificial-intelligence media python cross-platform comfyui voice-cloning voice-conversion rvc srt-subtitles multi-engine audio-editing node-extension audio gpu
1 source
- readme: https://github.com/diodiogod/TTS-Audio-Suite · fetched 2026-09-01 · 5cb7e6c78519
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| diodiogod/TTS-Audio-Suite | main | 84 |
For agents
markdown · JSON · MCP: product_card(name="diodiogod/TTS-Audio-Suite")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem