Ross ROSS = Recommend OSS · open-source software intelligence for agents

zzw922cn/awesome-speech-recognition-speech-synthesis-papers resource

Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC) observed · 2026-08-28

github.com/zzw922cn/awesome-speech-recognition-speech-synthesis-papers · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

32/100

  • Activity 0
  • Release rhythm 35
  • Longevity 100

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 3414
  • days_rel: n/a
  • days_push: 1049
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

3130 stars · 514 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A curated awesome-list of research papers on speech recognition, speech synthesis, speaker verification, voice conversion, and language modelling, spanning classic HMM work to modern diffusion-based audio generation. It serves as a reading roadmap rather than runnable software.

Use cases

  • find papers on automatic speech recognition
  • learn about text-to-speech research history
  • get a roadmap for studying speech synthesis
  • find voice conversion and speaker verification papers
  • discover text-to-audio and music generation papers
  • survey deep learning approaches to speech

When to choose

  • you need a curated paper list covering ASR and TTS from the 1980s to today
  • you are a researcher or student building a reading plan for speech processing
  • you want references spanning classic HMM methods to diffusion models

When to avoid

  • you need runnable speech recognition or TTS code
  • you want a maintained library or tool rather than a paper index
  • you need tutorials with code examples instead of paper links

Facets

learning-resource · maturity maintenance

speech-recognition tts nlp audio-processing documentation speech-processing tutorials artificial-intelligence cross-platform awesome-list papers asr voice-conversion speaker-verification singing-voice-synthesis deep-learning audio

1 source

Member repositories

For agents

markdown · JSON · MCP: product_card(name="zzw922cn/awesome-speech-recognition-speech-synthesis-papers")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem