nazdridoy/kokoro-tts
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents. observed · 2026-08-28
Health v2 · maintenance only
88/100
- Activity 99
- Release rhythm 99
- Longevity 42
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 26.0
- age_days: 596
- days_rel: 11
- days_push: 11
- n_releases_24m: 11
Adoption not part of the score
1816 stars · 174 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A Python CLI text-to-speech tool built on the Kokoro-82M model that converts text, EPUB, PDF, and TXT inputs into natural-sounding speech with multiple languages, voices, and voice blending. It supports WAV/MP3 output, streaming playback, chapter splitting/merging, adjustable speed, and GPU acceleration.
Use cases
- convert epub book to audiobook
- generate speech audio from a pdf document
- text to speech from the command line
- create mp3 narration of text files
- blend multiple tts voices with custom weights
- run offline text-to-speech with gpu acceleration
- split generated audio into chapters
When to choose
- you want a lightweight, fast, Apache-licensed TTS model runnable locally
- you need to convert ebooks or PDFs into audio files
- you prefer a terminal workflow with piping and stdin support
- you want voice blending and multi-language support without cloud APIs
When to avoid
- you need a graphical interface (GUI is still a TODO)
- you run Python 3.13+ (only 3.11-3.12 supported)
- you need real-time streaming synthesis at scale in production services
- you need languages or voices beyond what Kokoro-82M provides
Facets
cli-tool · maturity active
tts cli audio-processing pdf gpu-computing speech-processing artificial-intelligence pdf cli python cross-platform text-to-speech kokoro epub audiobook-generation voice-blending offline audio command-line gpu
8 sources
- readme: https://github.com/nazdridoy/kokoro-tts · fetched 2026-08-28 · 7416feae6370
- homepage: https://huggingface.co/hexgrad/Kokoro-82M · fetched 2026-08-29 · f9bbe8638b72
- site_page: https://huggingface.co/docs · fetched 2026-08-29 · bdec26667b98
- site_page: https://huggingface.co/docs/inference-providers · fetched 2026-08-29 · 8a5d0f819473
- site_page: https://huggingface.co/docs/hub/model-cards · fetched 2026-08-29 · 60ded09a56b0
- registry_pypi: https://pypi.org/pypi/kokoro-tts/json · fetched 2026-08-29 · 8442a87d45b7
- site_page: https://huggingface.co/pricing · fetched 2026-08-29 · de6b7a178be5
- site_page: https://huggingface.co/huggingface · fetched 2026-08-29 · a64a0fe552e5
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| nazdridoy/kokoro-tts | main | 88 |
For agents
markdown · JSON · MCP: product_card(name="nazdridoy/kokoro-tts")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem