WhiskeyCoder/Qwen3-Audiobook-Converter
Convert PDFs, EPUBs, DOCX, DOC, and TXT files into high-quality audiobooks using **Qwen3 TTS Voice Model** - an open-source voice synthesis system that excels at natural speech generation and voice cloning. observed · 2026-08-28
Health v2 · maintenance only
49/100
- Activity 76
- Release rhythm 35
- Longevity 15
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 221
- days_rel: n/a
- days_push: 148
- n_releases_24m: 0
Adoption not part of the score
1081 stars · 131 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A Python CLI tool that converts documents (PDF, EPUB, DOCX, DOC, TXT) into audiobooks using the Qwen3 TTS voice model running locally via a Gradio interface. It supports pre-built narration voices and voice cloning from reference audio samples, with smart text chunking, caching, and FFmpeg-based audio processing.
Use cases
- convert pdf books to audiobooks
- turn epub files into spoken audio
- clone a voice for audiobook narration
- batch convert docx documents to mp3 audio
- generate audiobooks with a local tts model
- make text-to-speech audiobooks offline
When to choose
- You want to convert written documents into audiobook audio files with natural-sounding narration
- You need voice cloning from a reference audio sample with automatic transcription
- You already run or are willing to run the Qwen3 TTS model locally and have FFmpeg installed
- You want a free, open-source alternative to commercial audiobook narration services
When to avoid
- You need a fully self-contained tool with no external TTS server dependency
- You want cloud-based or API-driven text-to-speech without local GPU/model setup
- You need formats beyond PDF, EPUB, DOCX, DOC, or TXT
- You require a graphical user interface rather than a command-line workflow
Facets
cli-tool · maturity active
tts audio-processing pdf speech-recognition cli pdf files media python cli cross-platform audiobook-converter text-to-speech voice-cloning qwen3-tts epub docx document-to-audio gradio ffmpeg audio natural-language-processing desktop
1 source
- readme: https://github.com/WhiskeyCoder/Qwen3-Audiobook-Converter · fetched 2026-08-28 · 10fc9c848b3c
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| WhiskeyCoder/Qwen3-Audiobook-Converter | main | 49 |
For agents
markdown · JSON · MCP: product_card(name="WhiskeyCoder/Qwen3-Audiobook-Converter")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem