Ross ROSS = Recommend OSS · open-source software intelligence for agents

FireRedTeam/FireRedTTS2

Long-form streaming TTS system for multi-speaker dialogue generation observed · 2026-08-28

github.com/FireRedTeam/FireRedTTS2 · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

40/100

  • Activity 49
  • Release rhythm 35
  • Longevity 26

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 365
  • days_rel: n/a
  • days_push: 311
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1428 stars · 128 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

FireRedTTS-2 is a long-form streaming text-to-speech system for multi-speaker dialogue generation, built in PyTorch with a dual-transformer architecture on a 12.5Hz streaming speech tokenizer. It supports zero-shot voice cloning, multilingual synthesis, and low first-packet latency (~140ms on an L20 GPU).

Use cases

  • generate podcast audio with multiple speakers from a script
  • build a chatbot with natural streaming voice responses
  • clone a voice with zero-shot samples for cross-lingual speech
  • create synthetic ASR training data with random timbres
  • generate long multi-speaker dialogues in Chinese, English, Japanese, Korean, French, German, or Russian

When to choose

  • you need low-latency streaming TTS for conversational agents or chatbots
  • you want multi-speaker dialogue generation with stable speaker switching
  • you need zero-shot voice cloning across multiple languages
  • you want to synthesize long-form podcast-style audio

When to avoid

  • you need a lightweight CPU-only TTS for embedded devices
  • you only need simple single-phrase TTS without conversational context
  • you require non-PyTorch deployment runtimes like ONNX or TensorRT out of the box

Facets

library · maturity active

tts speech-recognition machine-learning llm-inference speech-processing artificial-intelligence python voice-cloning podcast-generation streaming-tts multi-speaker-dialogue zero-shot-voice-cloning pytorch audio natural-language-processing gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
FireRedTeam/FireRedTTS2main40

For agents

markdown · JSON · MCP: product_card(name="FireRedTeam/FireRedTTS2")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem