Ross ROSS = Recommend OSS · open-source software intelligence for agents

lenML/Speech-AI-Forge

🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI. observed · 2026-08-28

github.com/lenML/Speech-AI-Forge · Python · AGPL-3.0 (copyleft) observed · 2026-08-28

Health v2 · maintenance only

62/100

  • Activity 83
  • Release rhythm 36
  • Longevity 58
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 823
  • days_rel: 212
  • days_push: 104
  • n_releases_24m: 1

Full methodology

Adoption not part of the score

1416 stars · 189 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Speech-AI-Forge is a Python project built around multiple TTS generation models (ChatTTS, CosyVoice, Fish-Speech, Index-TTS, F5-TTS, FireRedTTS, and more) that provides both an API server and a Gradio-based WebUI. It also supports ASR/STT via Whisper and SenseVoice, SSML, and cloud TTS backends like MiniMax.

Use cases

  • generate speech from text with multiple tts models
  • self-host a text-to-speech api server
  • clone a voice for tts synthesis
  • transcribe audio to text with whisper or sensevoice
  • run a gradio webui for text-to-speech
  • use ssml to control speech synthesis
  • try tts models in colab without installing

When to choose

  • you want one self-hosted server exposing many open-source TTS models behind a unified API
  • you need both a WebUI for experimentation and an API for integration
  • you want SSML support and voice cloning across multiple TTS engines
  • you need Chinese and English speech synthesis

When to avoid

  • you need a lightweight production TTS service with a single optimized model
  • you require a permissively licensed dependency since it is AGPL-3.0
  • you need real-time low-latency streaming TTS at scale
  • you only want cloud TTS without local model support

Facets

application · maturity active

tts speech-recognition http-server llm-inference api-framework speech-processing artificial-intelligence python windows cross-platform text-to-speech gradio-webui chattts cosyvoice fish-speech ssml voice-cloning asr whisper colab natural-language-processing audio docker web-server gpu

1 source

Member repositories

RepositoryRoleHealth v2
lenML/Speech-AI-Forgemain62

For agents

markdown · JSON · MCP: product_card(name="lenML/Speech-AI-Forge")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem