Ross ROSS = Recommend OSS · open-source software intelligence for agents

WhiskeyCoder/Qwen3-Audiobook-Converter

Convert PDFs, EPUBs, DOCX, DOC, and TXT files into high-quality audiobooks using **Qwen3 TTS Voice Model** - an open-source voice synthesis system that excels at natural speech generation and voice cloning. observed · 2026-08-28

github.com/WhiskeyCoder/Qwen3-Audiobook-Converter · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

49/100

  • Activity 76
  • Release rhythm 35
  • Longevity 15

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 221
  • days_rel: n/a
  • days_push: 148
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1081 stars · 131 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A Python CLI tool that converts documents (PDF, EPUB, DOCX, DOC, TXT) into audiobooks using the Qwen3 TTS voice model running locally via a Gradio interface. It supports pre-built narration voices and voice cloning from reference audio samples, with smart text chunking, caching, and FFmpeg-based audio processing.

Use cases

  • convert pdf books to audiobooks
  • turn epub files into spoken audio
  • clone a voice for audiobook narration
  • batch convert docx documents to mp3 audio
  • generate audiobooks with a local tts model
  • make text-to-speech audiobooks offline

When to choose

  • You want to convert written documents into audiobook audio files with natural-sounding narration
  • You need voice cloning from a reference audio sample with automatic transcription
  • You already run or are willing to run the Qwen3 TTS model locally and have FFmpeg installed
  • You want a free, open-source alternative to commercial audiobook narration services

When to avoid

  • You need a fully self-contained tool with no external TTS server dependency
  • You want cloud-based or API-driven text-to-speech without local GPU/model setup
  • You need formats beyond PDF, EPUB, DOCX, DOC, or TXT
  • You require a graphical user interface rather than a command-line workflow

Facets

cli-tool · maturity active

tts audio-processing pdf speech-recognition cli pdf files media python cli cross-platform audiobook-converter text-to-speech voice-cloning qwen3-tts epub docx document-to-audio gradio ffmpeg audio natural-language-processing desktop

1 source

Member repositories

RepositoryRoleHealth v2
WhiskeyCoder/Qwen3-Audiobook-Convertermain49

For agents

markdown · JSON · MCP: product_card(name="WhiskeyCoder/Qwen3-Audiobook-Converter")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem