Enemyx-net/VibeVoice-ComfyUI
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows. observed · 2026-08-28
Health v2 · maintenance only
53/100
- Activity 68
- Release rhythm 50
- Longevity 26
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 0
- age_days: 371
- days_rel: 335
- days_push: 196
- n_releases_24m: 30
Adoption not part of the score
1549 stars · 249 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A ComfyUI custom node integration for Microsoft's VibeVoice text-to-speech model, providing single and multi-speaker voice synthesis with voice cloning inside ComfyUI workflows. It supports LoRA voice fine-tuning, quantization for lower VRAM, and CUDA, CPU, and Apple Silicon MPS backends.
Use cases
- generate speech from text in comfyui workflows
- clone a voice from an audio sample
- create multi-speaker dialogue audio with up to 4 speakers
- narrate videos or audiobooks with tts nodes
- run tts locally on apple silicon or low-vram gpus
- fine-tune voices with custom lora adapters
- insert pauses and control speech speed in generated audio
When to choose
- you already use ComfyUI and want TTS or voice cloning as nodes in your workflows
- you need multi-speaker conversation synthesis or LoRA-based voice customization
- you want local, self-hosted TTS with CUDA or Apple Silicon acceleration
When to avoid
- you need a standalone TTS service or CLI outside ComfyUI
- you want fully automatic model downloads without manual setup
- you need real-time streaming TTS for live applications
Facets
plugin · maturity active
tts audio-processing machine-learning llm-inference artificial-intelligence speech-processing media windows python cross-platform comfyui comfyui-custom-node voice-cloning multi-speaker-tts vibevoice text-to-speech ai-audio lora-support quantization audio linux macos gpu
1 source
- readme: https://github.com/Enemyx-net/VibeVoice-ComfyUI · fetched 2026-08-28 · f4072a7e17ce
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| Enemyx-net/VibeVoice-ComfyUI | main | 53 |
For agents
markdown · JSON · MCP: product_card(name="Enemyx-net/VibeVoice-ComfyUI")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem