# PlayVoice/whisper-vits-svc

Core Engine of Singing Voice Conversion & Singing Voice Clone

Repository: https://github.com/PlayVoice/whisper-vits-svc
Canonical: https://ross.abutalabs.com/products/whisper-vits-svc
Homepage: https://huggingface.co/spaces/maxmax20160403/sovits5.0
Language: Python
License: MIT
License Family: permissive
Topics: sovits, svc, vits, change, voice, singing-voice-conversion, diff-svc, diffusion, diffusion-svc, vits2
Last push: 2024-04-23T10:50:32+00:00

## Health v2 (maintenance only)
Score: 23/100 (v2, computed 2026-09-03T02:39:23.370411+00:00)
- activity 0, release rhythm 8, longevity 100
- inputs: {"age_days": 1442, "days_push": 862, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: none
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 2864, forks 907 (observed 2026-08-28T04:07:26.440690+00:00)

## What it is
A PyTorch-based singing voice conversion and voice cloning engine built on VITS with Whisper, BigVGAN, and diffusion components. It lets users train multi-speaker SVC models and convert singing voices, including with light accompaniment.

## Use cases
- convert a singing voice to another singer's timbre
- clone a singing voice from audio samples
- train a so-vits-svc model on my own dataset
- change the voice of a song while keeping the accompaniment
- create a multi-speaker singing voice conversion model
- edit pitch (F0) of converted singing audio

## When to choose
- you want to train or fine-tune a singing voice conversion model in Python/PyTorch
- you need multi-speaker SVC or one-shot voice cloning
- you have a GPU with at least 6GB VRAM and want an open-source SVC engine

## When to avoid
- you need real-time voice conversion
- you want a one-click GUI package with no coding
- you need active development or new features - the project is no longer upgraded

## Facets
- artifact type: library
- maturity: maintenance
- function: machine-learning, deep-learning, audio-processing, speech-recognition, tts
- domain: deep-learning, machine-learning, artificial-intelligence
- platform: python, cross-platform
- tags: singing-voice-conversion, voice-cloning, so-vits-svc, vits, diffusion, whisper, bigvgan, pytorch, audio, gpu

## Member repositories
- PlayVoice/whisper-vits-svc (main) score 23

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:07:26.440690+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T07:36:14.655190+00:00, confidence not recorded.
  - readme: https://github.com/PlayVoice/whisper-vits-svc (fetched 2026-08-28T04:07:26.440690+00:00, sha b09419a6b813)
  - homepage: https://huggingface.co/spaces/maxmax20160403/sovits5.0 (fetched 2026-08-29T09:51:51.822155+00:00, sha 0f5e0cd76c89)
- Data as of 2026-08-30T08:39:29.467469+00:00.
