Ross ROSS = Recommend OSS · open-source software intelligence for agents

readbeyond/aeneas

aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment) observed · 2026-08-28

github.com/readbeyond/aeneas · homepage · Python · AGPL-3.0 (copyleft) observed · 2026-08-28

Health v2 · maintenance only

65/100

  • Activity 94
  • Release rhythm 8
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 4132
  • days_rel: n/a
  • days_push: 39
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

2863 stars · 275 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

aeneas is a Python/C library and set of CLI tools that automatically computes forced alignments, generating a synchronization map between a list of text fragments and an audio file narrating that text. It supports many output formats including SRT, SMIL (EPUB 3 Media Overlays), ELAN EAF, TextGrid, and JSON, using MFCC features and dynamic time warping with eSpeak/eSpeak-ng or Festival TTS.

Use cases

  • align audiobook narration with text fragments
  • generate subtitles from a transcript and audio
  • create EPUB 3 media overlay SMIL files for audio-eBooks
  • synchronize lyrics or poetry with a recording
  • produce ELAN or TextGrid annotations for speech research
  • batch-align many audio/text pairs via job files

When to choose

  • you need to time-align a known transcript to its audio without manual annotation
  • you want EPUB 3 audio-eBook synchronization maps
  • you need subtitle files (SRT, VTT, TTML) generated from text and narration
  • you work offline with a lightweight Python/C tool rather than heavy deep-learning ASR

When to avoid

  • you need state-of-the-art speech recognition or alignment on noisy, spontaneous audio
  • you want an actively developed project with recent releases and modern Python support
  • you need word-level alignment with neural acoustic models
  • you cannot install C dependencies like ffmpeg and espeak-ng

Facets

library · maturity maintenance

speech-recognition tts audio-processing nlp cli speech-processing media accessibility python windows cli forced-alignment dtw sync-map subtitles epub3-media-overlays espeak ffmpeg audio-text-synchronization audio natural-language-processing linux macos

10 sources

Member repositories

RepositoryRoleHealth v2
readbeyond/aeneasmain65

For agents

markdown · JSON · MCP: product_card(name="readbeyond/aeneas")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem