erew123/alltalk_tts
AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advanced features, such as a settings page, low VRAM support, DeepSpeed, narrator, model finetuning, custom models, wav file maintenance. It can also be used with 3rd Party software via JSON calls. observed · 2026-08-28
Health v2 · maintenance only
44/100
- Activity 61
- Release rhythm 8
- Longevity 71
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 999
- days_rel: 555
- days_push: 236
- n_releases_24m: 1
Adoption not part of the score
2429 stars · 285 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
AllTalk TTS is a text-to-speech application built on the Coqui TTS engine, usable standalone or as an extension for Text-generation-webui, SillyTavern, and KoboldCPP. It offers advanced features like model finetuning, DeepSpeed acceleration, low VRAM mode, multi-voice narration, and a JSON API for third-party integration.
Use cases
- generate speech from text locally with a gpu
- add tts voices to sillytavern or text-generation-webui
- clone a custom voice and finetune a tts model
- generate audiobooks with separate narrator and character voices
- run text-to-speech on a low vram gpu alongside an llm
- integrate tts into third-party apps via a json api
When to choose
- you want local, self-hosted text-to-speech with xttsv2 models
- you need voice finetuning or multi-voice narration
- you use text-generation-webui, SillyTavern, or KoboldCPP and want tts integration
- you have limited vram or want DeepSpeed speedups
When to avoid
- you need a managed cloud tts service
- you have no gpu and need very fast synthesis
- you need a lightweight library to embed rather than a full application
Facets
application · maturity active
tts speech-recognition llm-inference api-framework audio-processing speech-processing artificial-intelligence large-language-models windows python self-hosted coqui-tts xttsv2 text-to-speech deepspeed sillytavern text-generation-webui voice-cloning narration audio linux gpu
1 source
- readme: https://github.com/erew123/alltalk_tts · fetched 2026-08-28 · 1318cb3d75ce
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| erew123/alltalk_tts | main | 44 |
For agents
markdown · JSON · MCP: product_card(name="erew123/alltalk_tts")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem