Ross ROSS = Recommend OSS · open-source software intelligence for agents

diodiogod/TTS-Audio-Suite

A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools observed · 2026-09-01

github.com/diodiogod/TTS-Audio-Suite · Python · NOASSERTION (other) observed · 2026-09-01

Health v2 · maintenance only

84/100

  • Activity 100
  • Release rhythm 95
  • Longevity 28

Flags: no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 7
  • age_days: 392
  • days_rel: 39
  • days_push: 2
  • n_releases_24m: 22

Full methodology

Adoption not part of the score

1185 stars · 141 forks observed · 2026-09-01

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A ComfyUI custom node suite providing unified multi-engine Text-to-Speech, Voice Conversion, and audio editing across 19 engines like ChatterBox, F5-TTS, Higgs Audio, and RVC. It also supports SRT subtitle workflows, character-based dialogue, and integrated RVC model training.

Use cases

  • generate speech from text locally with multiple TTS engines
  • clone a voice from a reference audio sample
  • convert one voice to another with RVC
  • generate character dialogue for videos with SRT timing
  • transcribe audio to subtitles and rebuild them from edited transcripts
  • edit speech emotion and style in existing audio
  • build TTS workflows inside ComfyUI

When to choose

  • you already use ComfyUI and want TTS or voice conversion nodes
  • you need to compare or combine multiple TTS engines in one tool
  • you need subtitle-synced, multi-character voiceover generation
  • you want local, offline speech generation with voice cloning

When to avoid

  • you need a simple standalone TTS CLI without ComfyUI
  • you lack a GPU or the disk space for multi-GB models
  • you need production cloud TTS with SLAs
  • you want a lightweight plugin with minimal dependencies

Facets

plugin · maturity active

tts audio-processing speech-recognition machine-learning plugin-system speech-processing artificial-intelligence media python cross-platform comfyui voice-cloning voice-conversion rvc srt-subtitles multi-engine audio-editing node-extension audio gpu

1 source

Member repositories

RepositoryRoleHealth v2
diodiogod/TTS-Audio-Suitemain84

For agents

markdown · JSON · MCP: product_card(name="diodiogod/TTS-Audio-Suite")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem