readbeyond/aeneas
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment) observed · 2026-08-28
Health v2 · maintenance only
65/100
- Activity 94
- Release rhythm 8
- Longevity 100
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 4132
- days_rel: n/a
- days_push: 39
- n_releases_24m: 0
Adoption not part of the score
2863 stars · 275 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
aeneas is a Python/C library and set of CLI tools that automatically computes forced alignments, generating a synchronization map between a list of text fragments and an audio file narrating that text. It supports many output formats including SRT, SMIL (EPUB 3 Media Overlays), ELAN EAF, TextGrid, and JSON, using MFCC features and dynamic time warping with eSpeak/eSpeak-ng or Festival TTS.
Use cases
- align audiobook narration with text fragments
- generate subtitles from a transcript and audio
- create EPUB 3 media overlay SMIL files for audio-eBooks
- synchronize lyrics or poetry with a recording
- produce ELAN or TextGrid annotations for speech research
- batch-align many audio/text pairs via job files
When to choose
- you need to time-align a known transcript to its audio without manual annotation
- you want EPUB 3 audio-eBook synchronization maps
- you need subtitle files (SRT, VTT, TTML) generated from text and narration
- you work offline with a lightweight Python/C tool rather than heavy deep-learning ASR
When to avoid
- you need state-of-the-art speech recognition or alignment on noisy, spontaneous audio
- you want an actively developed project with recent releases and modern Python support
- you need word-level alignment with neural acoustic models
- you cannot install C dependencies like ffmpeg and espeak-ng
Facets
library · maturity maintenance
speech-recognition tts audio-processing nlp cli speech-processing media accessibility python windows cli forced-alignment dtw sync-map subtitles epub3-media-overlays espeak ffmpeg audio-text-synchronization audio natural-language-processing linux macos
10 sources
- readme: https://github.com/readbeyond/aeneas · fetched 2026-08-28 · e27199e75fbc
- homepage: http://www.readbeyond.it/aeneas/ · fetched 2026-08-29 · b92ff4d76b84
- site_page: http://www.readbeyond.it/aeneas/docs/libtutorial.html · fetched 2026-08-29 · f41706419ffa
- site_page: http://www.readbeyond.it/aeneas/docs/changelog.html · fetched 2026-08-29 · 12fa9cb3fbaf
- site_page: https://www.readbeyond.it/about.html · fetched 2026-08-29 · 22e1090e98dc
- site_page: https://www.readbeyond.it/aeneas/docs · fetched 2026-08-29 · ddb5da610293
- site_page: http://www.readbeyond.it/aeneas/docs · fetched 2026-08-29 · ddb5da610293
- site_page: http://www.readbeyond.it/aeneas/docs/clitutorial.html · fetched 2026-08-29 · e72ba305471c
- site_page: http://www.readbeyond.it/aeneas/docs/textfile.html · fetched 2026-08-29 · fa51172bc4a4
- registry_pypi: https://pypi.org/pypi/aeneas/json · fetched 2026-08-29 · 116b5a823dd4
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| readbeyond/aeneas | main | 65 |
For agents
markdown · JSON · MCP: product_card(name="readbeyond/aeneas")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem