# Edresson/YourTTS

YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

Repository: https://github.com/Edresson/YourTTS
Canonical: https://ross.abutalabs.com/products/yourtts
Language: Jupyter Notebook
License: NOASSERTION
License Family: other
Topics: voice-conversion, zero-shot-voice-conversion, zero-shot-multi-speaker-tts, tts, speech-synthesis
Last push: 2024-11-04T12:00:41+00:00

## Health v2 (maintenance only)
Score: 23/100 (v2, computed 2026-09-02T17:46:02.011165+00:00)
- activity 0, release rhythm 8, longevity 100
- inputs: {"age_days": 1792, "days_push": 667, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_license
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1053, forks 97 (observed 2026-08-28T04:03:23.821188+00:00)

## What it is
YourTTS is a zero-shot multi-speaker text-to-speech and voice conversion model built on VITS, implemented in the Coqui TTS framework. It supports multilingual synthesis, voice cloning from short reference clips, and fine-tuning with under a minute of speech.

## Use cases
- clone a voice from a short audio sample
- synthesize speech in a target speaker's voice from text
- convert speech from one voice to another
- build TTS for low-resource languages
- fine-tune a TTS model with less than a minute of data

## When to choose
- you need zero-shot multi-speaker TTS or voice conversion
- you want to clone voices from minimal reference audio
- you work with low-resource languages and need multilingual TTS
- you want a research-grade model integrated with Coqui TTS

## When to avoid
- you need a production-ready, actively maintained TTS pipeline
- you want a simple plug-and-play API without ML setup
- you need fullband commercial-quality English-only synthesis
- you require a permissively licensed model

## Facets
- artifact type: library
- maturity: maintenance
- function: tts, speech-recognition, machine-learning, deep-learning
- domain: speech-processing, machine-learning, artificial-intelligence
- platform: python, cross-platform
- tags: zero-shot-tts, voice-conversion, vits, speech-synthesis, voice-cloning, research-model, coqui-tts, gpu

## Member repositories
- Edresson/YourTTS (main) score 23

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:03:23.821188+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T06:59:18.858494+00:00, confidence not recorded.
  - readme: https://github.com/Edresson/YourTTS (fetched 2026-08-28T04:03:23.821188+00:00, sha 26e6b5894ca5)
- Data as of 2026-08-30T08:39:29.467469+00:00.
