Ross ROSS = Recommend OSS · open-source software intelligence for agents

Enemyx-net/VibeVoice-ComfyUI

A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows. observed · 2026-08-28

github.com/Enemyx-net/VibeVoice-ComfyUI · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

53/100

  • Activity 68
  • Release rhythm 50
  • Longevity 26
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 0
  • age_days: 371
  • days_rel: 335
  • days_push: 196
  • n_releases_24m: 30

Full methodology

Adoption not part of the score

1549 stars · 249 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A ComfyUI custom node integration for Microsoft's VibeVoice text-to-speech model, providing single and multi-speaker voice synthesis with voice cloning inside ComfyUI workflows. It supports LoRA voice fine-tuning, quantization for lower VRAM, and CUDA, CPU, and Apple Silicon MPS backends.

Use cases

  • generate speech from text in comfyui workflows
  • clone a voice from an audio sample
  • create multi-speaker dialogue audio with up to 4 speakers
  • narrate videos or audiobooks with tts nodes
  • run tts locally on apple silicon or low-vram gpus
  • fine-tune voices with custom lora adapters
  • insert pauses and control speech speed in generated audio

When to choose

  • you already use ComfyUI and want TTS or voice cloning as nodes in your workflows
  • you need multi-speaker conversation synthesis or LoRA-based voice customization
  • you want local, self-hosted TTS with CUDA or Apple Silicon acceleration

When to avoid

  • you need a standalone TTS service or CLI outside ComfyUI
  • you want fully automatic model downloads without manual setup
  • you need real-time streaming TTS for live applications

Facets

plugin · maturity active

tts audio-processing machine-learning llm-inference artificial-intelligence speech-processing media windows python cross-platform comfyui comfyui-custom-node voice-cloning multi-speaker-tts vibevoice text-to-speech ai-audio lora-support quantization audio linux macos gpu

1 source

Member repositories

RepositoryRoleHealth v2
Enemyx-net/VibeVoice-ComfyUImain53

For agents

markdown · JSON · MCP: product_card(name="Enemyx-net/VibeVoice-ComfyUI")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem