Ross ROSS = Recommend OSS · open-source software intelligence for agents

Aratako/Irodori-TTS

A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control observed · 2026-08-28

github.com/Aratako/Irodori-TTS · homepage · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

58/100

  • Activity 97
  • Release rhythm 35
  • Longevity 13

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 189
  • days_rel: n/a
  • days_push: 23
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1219 stars · 151 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Irodori-TTS is a Flow Matching-based text-to-speech model with training and inference code, built on a Rectified Flow Diffusion Transformer over DACVAE latents. It supports zero-shot voice cloning, multi-modal voice design with emoji-based style control, LoRA fine-tuning, and CLI/Gradio inference.

Use cases

  • clone a voice from reference audio
  • generate speech with emotion and style control
  • design a custom voice from text descriptions
  • fine-tune a TTS model with LoRA on my own voice
  • train a flow matching text-to-speech model
  • generate japanese speech from text
  • run tts inference from the command line

When to choose

  • you need zero-shot voice cloning or voice design with style/emotion control
  • you want to fine-tune or train a modern flow-matching TTS model
  • you want emoji-annotated text to influence speech delivery
  • you need a compact open TTS model with released checkpoints and demos

When to avoid

  • you need a production-ready OpenAI-compatible serving API out of the box (use the companion Irodori-TTS-Server)
  • you need non-Japanese or multilingual TTS
  • you have no GPU and need fast low-latency synthesis
  • you just want a plug-and-play TTS without touching model code

Facets

library · maturity active

tts machine-learning deep-learning audio-processing llm-training speech-processing machine-learning python cli cross-platform flow-matching diffusion-transformer voice-cloning voice-design emoji-style-control dacvae lora-fine-tuning japanese-tts audio-watermarking gradio audio natural-language-processing gpu

6 sources

Member repositories

RepositoryRoleHealth v2
Aratako/Irodori-TTSmain58

For agents

markdown · JSON · MCP: product_card(name="Aratako/Irodori-TTS")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem