Ross ROSS = Recommend OSS · open-source software intelligence for agents

erew123/alltalk_tts

AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advanced features, such as a settings page, low VRAM support, DeepSpeed, narrator, model finetuning, custom models, wav file maintenance. It can also be used with 3rd Party software via JSON calls. observed · 2026-08-28

github.com/erew123/alltalk_tts · HTML · AGPL-3.0 (copyleft) observed · 2026-08-28

Health v2 · maintenance only

44/100

  • Activity 61
  • Release rhythm 8
  • Longevity 71
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 999
  • days_rel: 555
  • days_push: 236
  • n_releases_24m: 1

Full methodology

Adoption not part of the score

2429 stars · 285 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

AllTalk TTS is a text-to-speech application built on the Coqui TTS engine, usable standalone or as an extension for Text-generation-webui, SillyTavern, and KoboldCPP. It offers advanced features like model finetuning, DeepSpeed acceleration, low VRAM mode, multi-voice narration, and a JSON API for third-party integration.

Use cases

  • generate speech from text locally with a gpu
  • add tts voices to sillytavern or text-generation-webui
  • clone a custom voice and finetune a tts model
  • generate audiobooks with separate narrator and character voices
  • run text-to-speech on a low vram gpu alongside an llm
  • integrate tts into third-party apps via a json api

When to choose

  • you want local, self-hosted text-to-speech with xttsv2 models
  • you need voice finetuning or multi-voice narration
  • you use text-generation-webui, SillyTavern, or KoboldCPP and want tts integration
  • you have limited vram or want DeepSpeed speedups

When to avoid

  • you need a managed cloud tts service
  • you have no gpu and need very fast synthesis
  • you need a lightweight library to embed rather than a full application

Facets

application · maturity active

tts speech-recognition llm-inference api-framework audio-processing speech-processing artificial-intelligence large-language-models windows python self-hosted coqui-tts xttsv2 text-to-speech deepspeed sillytavern text-generation-webui voice-cloning narration audio linux gpu

1 source

Member repositories

RepositoryRoleHealth v2
erew123/alltalk_ttsmain44

For agents

markdown · JSON · MCP: product_card(name="erew123/alltalk_tts")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem