Ross ROSS = Recommend OSS · open-source software intelligence for agents

nazdridoy/kokoro-tts

A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents. observed · 2026-08-28

github.com/nazdridoy/kokoro-tts · homepage · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

88/100

  • Activity 99
  • Release rhythm 99
  • Longevity 42
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 26.0
  • age_days: 596
  • days_rel: 11
  • days_push: 11
  • n_releases_24m: 11

Full methodology

Adoption not part of the score

1816 stars · 174 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A Python CLI text-to-speech tool built on the Kokoro-82M model that converts text, EPUB, PDF, and TXT inputs into natural-sounding speech with multiple languages, voices, and voice blending. It supports WAV/MP3 output, streaming playback, chapter splitting/merging, adjustable speed, and GPU acceleration.

Use cases

  • convert epub book to audiobook
  • generate speech audio from a pdf document
  • text to speech from the command line
  • create mp3 narration of text files
  • blend multiple tts voices with custom weights
  • run offline text-to-speech with gpu acceleration
  • split generated audio into chapters

When to choose

  • you want a lightweight, fast, Apache-licensed TTS model runnable locally
  • you need to convert ebooks or PDFs into audio files
  • you prefer a terminal workflow with piping and stdin support
  • you want voice blending and multi-language support without cloud APIs

When to avoid

  • you need a graphical interface (GUI is still a TODO)
  • you run Python 3.13+ (only 3.11-3.12 supported)
  • you need real-time streaming synthesis at scale in production services
  • you need languages or voices beyond what Kokoro-82M provides

Facets

cli-tool · maturity active

tts cli audio-processing pdf gpu-computing speech-processing artificial-intelligence pdf cli python cross-platform text-to-speech kokoro epub audiobook-generation voice-blending offline audio command-line gpu

8 sources

Member repositories

RepositoryRoleHealth v2
nazdridoy/kokoro-ttsmain88

For agents

markdown · JSON · MCP: product_card(name="nazdridoy/kokoro-tts")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem