jamiepine/voicebox
The open-source AI voice studio. Clone, dictate, create. observed · 2026-08-28
Health v2 · maintenance only
75/100
- Activity 96
- Release rhythm 81
- Longevity 15
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 0.0
- age_days: 220
- days_rel: 130
- days_push: 25
- n_releases_24m: 25
Adoption not part of the score
51551 stars · 6430 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
Voicebox is a free, open-source, local-first AI voice studio desktop app that combines voice cloning, text-to-speech across 7 engines, and system-wide voice dictation. It runs entirely on your own hardware and can give MCP-aware AI agents a voice of your choosing.
Use cases
- clone a voice from a few seconds of audio
- generate speech locally instead of paying for ElevenLabs
- dictate into any app with a global hotkey instead of WisprFlow
- give my AI agent a custom voice via MCP
- generate multilingual TTS in 23 languages
- create multi-voice audio projects and stories
- transcribe and archive voice recordings locally
When to choose
- you want free, private, fully local voice cloning and TTS with no account or subscription
- you need both speech output (TTS) and input (dictation) in one desktop app
- you want to give MCP-aware AI agents a voice you own
- you have a GPU (CUDA or Apple Silicon) and want offline speech generation
When to avoid
- you need a managed cloud API with guaranteed uptime and no local hardware
- you need a prebuilt Linux binary — Linux currently requires building from source
- you only need lightweight command-line TTS scripting rather than a GUI studio
- you require commercial licensing guarantees for cloned celebrity-style voices
Facets
application · maturity active
tts speech-recognition llm-inference mcp audio-processing gui speech-processing artificial-intelligence desktop-applications self-hosted windows cross-platform voice-cloning dictation local-first tauri elevenlabs-alternative wisprflow-alternative qwen3-tts whisper voice-studio mcp-server audio macos linux desktop gpu
6 sources
- readme: https://github.com/jamiepine/voicebox · fetched 2026-08-28 · 7d37a432c73c
- homepage: https://voicebox.sh · fetched 2026-08-28 · 7b390bf104c5
- site_page: https://voicebox.sh/linux-install · fetched 2026-08-28 · 62d404dabb07
- site_page: https://docs.voicebox.sh · fetched 2026-08-28 · f6385afe3cdb
- site_page: https://voicebox.sh/pricing · fetched 2026-08-28 · 786a7fbe36b2
- site_page: https://voicebox.sh/token · fetched 2026-08-28 · a2cbf64af9ca
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| jamiepine/voicebox | main | 75 |
For agents
markdown · JSON · MCP: product_card(name="jamiepine/voicebox")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem