Aratako/Irodori-TTS
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control observed · 2026-08-28
Health v2 · maintenance only
58/100
- Activity 97
- Release rhythm 35
- Longevity 13
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 189
- days_rel: n/a
- days_push: 23
- n_releases_24m: 0
Adoption not part of the score
1219 stars · 151 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
Irodori-TTS is a Flow Matching-based text-to-speech model with training and inference code, built on a Rectified Flow Diffusion Transformer over DACVAE latents. It supports zero-shot voice cloning, multi-modal voice design with emoji-based style control, LoRA fine-tuning, and CLI/Gradio inference.
Use cases
- clone a voice from reference audio
- generate speech with emotion and style control
- design a custom voice from text descriptions
- fine-tune a TTS model with LoRA on my own voice
- train a flow matching text-to-speech model
- generate japanese speech from text
- run tts inference from the command line
When to choose
- you need zero-shot voice cloning or voice design with style/emotion control
- you want to fine-tune or train a modern flow-matching TTS model
- you want emoji-annotated text to influence speech delivery
- you need a compact open TTS model with released checkpoints and demos
When to avoid
- you need a production-ready OpenAI-compatible serving API out of the box (use the companion Irodori-TTS-Server)
- you need non-Japanese or multilingual TTS
- you have no GPU and need fast low-latency synthesis
- you just want a plug-and-play TTS without touching model code
Facets
library · maturity active
tts machine-learning deep-learning audio-processing llm-training speech-processing machine-learning python cli cross-platform flow-matching diffusion-transformer voice-cloning voice-design emoji-style-control dacvae lora-fine-tuning japanese-tts audio-watermarking gradio audio natural-language-processing gpu
6 sources
- readme: https://github.com/Aratako/Irodori-TTS · fetched 2026-08-28 · 67ea56689801
- homepage: https://huggingface.co/collections/Aratako/irodori-tts · fetched 2026-08-29 · c3d6002da11c
- site_page: https://huggingface.co/docs · fetched 2026-08-29 · bdec26667b98
- site_page: https://huggingface.co/docs/hub/collections · fetched 2026-08-29 · 8bc6746b9e69
- site_page: https://huggingface.co/pricing · fetched 2026-08-29 · de6b7a178be5
- site_page: https://huggingface.co/huggingface · fetched 2026-08-29 · 6ae0067a4ae4
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| Aratako/Irodori-TTS | main | 58 |
For agents
markdown · JSON · MCP: product_card(name="Aratako/Irodori-TTS")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem