Ross ROSS = Recommend OSS · open-source software intelligence for agents

jamiepine/voicebox

The open-source AI voice studio. Clone, dictate, create. observed · 2026-08-28

github.com/jamiepine/voicebox · homepage · TypeScript · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

75/100

  • Activity 96
  • Release rhythm 81
  • Longevity 15
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 0.0
  • age_days: 220
  • days_rel: 130
  • days_push: 25
  • n_releases_24m: 25

Full methodology

Adoption not part of the score

51551 stars · 6430 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

Voicebox is a free, open-source, local-first AI voice studio desktop app that combines voice cloning, text-to-speech across 7 engines, and system-wide voice dictation. It runs entirely on your own hardware and can give MCP-aware AI agents a voice of your choosing.

Use cases

  • clone a voice from a few seconds of audio
  • generate speech locally instead of paying for ElevenLabs
  • dictate into any app with a global hotkey instead of WisprFlow
  • give my AI agent a custom voice via MCP
  • generate multilingual TTS in 23 languages
  • create multi-voice audio projects and stories
  • transcribe and archive voice recordings locally

When to choose

  • you want free, private, fully local voice cloning and TTS with no account or subscription
  • you need both speech output (TTS) and input (dictation) in one desktop app
  • you want to give MCP-aware AI agents a voice you own
  • you have a GPU (CUDA or Apple Silicon) and want offline speech generation

When to avoid

  • you need a managed cloud API with guaranteed uptime and no local hardware
  • you need a prebuilt Linux binary — Linux currently requires building from source
  • you only need lightweight command-line TTS scripting rather than a GUI studio
  • you require commercial licensing guarantees for cloned celebrity-style voices

Facets

application · maturity active

tts speech-recognition llm-inference mcp audio-processing gui speech-processing artificial-intelligence desktop-applications self-hosted windows cross-platform voice-cloning dictation local-first tauri elevenlabs-alternative wisprflow-alternative qwen3-tts whisper voice-studio mcp-server audio macos linux desktop gpu

6 sources

Member repositories

RepositoryRoleHealth v2
jamiepine/voiceboxmain75

For agents

markdown · JSON · MCP: product_card(name="jamiepine/voicebox")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem