Ross ROSS = Recommend OSS · open-source software intelligence for agents

DigitalPhonetics/IMS-Toucan

Controllable and fast Text-to-Speech for over 7000 languages! observed · 2026-08-28

github.com/DigitalPhonetics/IMS-Toucan · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

63/100

  • Activity 64
  • Release rhythm 40
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: 14
  • age_days: 1854
  • days_rel: 695
  • days_push: 220
  • n_releases_24m: 2

Full methodology

Adoption not part of the score

2207 stars · 316 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

IMS Toucan is a PyTorch-based toolkit for training and running state-of-the-art, controllable text-to-speech synthesis, home of the massively multilingual ToucanTTS model supporting over 7000 languages. It is developed at the University of Stuttgart and designed to be fast and trainable without massive compute.

Use cases

  • synthesize speech in many languages including low-resource ones
  • train a custom TTS model on my own dataset
  • generate controllable speech with prosody and speaker control
  • run fast TTS inference without a GPU
  • build a multilingual voice assistant or audiobook pipeline
  • experiment with speech synthesis research models

When to choose

  • you need multilingual or low-resource language TTS coverage
  • you want to train or fine-tune TTS models on modest hardware
  • you need a Python toolkit with pretrained models and an Apache-2.0 license

When to avoid

  • you need production cloud TTS with managed APIs
  • you only need simple English TTS with minimal setup
  • you need real-time streaming TTS in non-Python environments

Facets

library · maturity active

tts speech-recognition machine-learning deep-learning audio-processing speech-processing machine-learning python windows text-to-speech speech-synthesis multilingual pytorch toucantts voice-cloning low-resource-languages natural-language-processing audio linux gpu

1 source

Member repositories

RepositoryRoleHealth v2
DigitalPhonetics/IMS-Toucanmain63

For agents

markdown · JSON · MCP: product_card(name="DigitalPhonetics/IMS-Toucan")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem