# RVC-Project/Retrieval-based-Voice-Conversion-WebUI

Easily train a good VC model with voice data <= 10 mins!

Repository: https://github.com/RVC-Project/Retrieval-based-Voice-Conversion-WebUI
Canonical: https://ross.abutalabs.com/products/retrieval-based-voice-conversion-webui
Language: Python
License: MIT
License Family: permissive
Topics: change, sovits, vits, voice, voice-conversion, rvc, audio-analysis, conversational-ai, conversion, converter, retrieval-model, retrieve-data, so-vits-svc, vc, voice-converter, voiceconversion
Last push: 2026-08-04T07:47:32+00:00

## Health v2 (maintenance only)
Score: 82/100 (v2, computed 2026-09-03T02:39:23.370411+00:00)
- activity 96, release rhythm 61, longevity 89
- inputs: {"age_days": 1255, "days_push": 29, "days_rel": 44, "gap_med": null, "n_releases_24m": 1}
- flags: none
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 37845, forks 5239 (observed 2026-08-28T04:12:01.595599+00:00)

## What it is
A WebUI-based framework for training and running retrieval-based voice conversion (RVC) models, letting users clone a voice timbre from as little as 10 minutes of clean speech. It includes training and inference interfaces, a real-time voice changer with ~90-170ms latency, and vocal/instrumental separation via MSS models.

## Use cases
- train an AI voice conversion model from 10 minutes of audio
- convert my voice to another voice in real time
- make an AI singing voice cover of a song
- clone a voice timbre from short recordings
- separate vocals from a song's instrumental
- run a low-latency voice changer for streaming or calls

## When to choose
- you want high-quality voice conversion trained on very little data
- you need a GUI for both training and real-time inference
- you want to avoid timbre leakage via retrieval-based feature replacement
- you have a modest GPU or even CPU/AMD/Intel fallback options

## When to avoid
- you need text-to-speech synthesis rather than voice-to-speech conversion
- you want a production API service rather than a local WebUI
- you have no GPU and need very fast batch inference
- you need fully automated headless pipelines with no GUI

## Facets
- artifact type: application
- maturity: active
- function: machine-learning, audio-processing, speech-recognition, llm-training, gui
- domain: machine-learning, deep-learning, media
- platform: python, windows, cross-platform
- tags: voice-conversion, voice-cloning, ai-singer, realtime-voice-changer, webui, so-vits-svc, vocal-separation, audio, linux, gpu

## Member repositories
- RVC-Project/Retrieval-based-Voice-Conversion-WebUI (main) score 82

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:12:01.595599+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-29T16:25:02.972141+00:00, confidence not recorded.
  - readme: https://github.com/RVC-Project/Retrieval-based-Voice-Conversion-WebUI (fetched 2026-08-28T04:12:01.595599+00:00, sha c09df9e4d60a)
- Data as of 2026-08-30T08:39:29.467469+00:00.
