Ross ROSS = Recommend OSS · open-source software intelligence for agents

speechbrain/speechbrain

A PyTorch-based Speech Toolkit observed · 2026-08-28

github.com/speechbrain/speechbrain · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

83/100

  • Activity 99
  • Release rhythm 53
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 93.5
  • age_days: 2318
  • days_rel: 156
  • days_push: 8
  • n_releases_24m: 5

Full methodology

Adoption not part of the score

11785 stars · 1718 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

SpeechBrain is an open-source PyTorch-based speech toolkit for building conversational AI systems. It provides training recipes, pretrained models, and tools for speech recognition, speaker recognition, speech enhancement/separation, TTS, and language modeling.

Use cases

  • transcribe speech to text with pretrained models
  • fine-tune whisper or wav2vec2 on my own dataset
  • verify speaker identity from voice recordings
  • separate overlapping speakers in an audio file
  • enhance noisy speech recordings
  • train a speech recognition model from scratch
  • build a voice assistant pipeline
  • diarize who spoke when in a meeting recording

When to choose

  • you want a flexible PyTorch-based toolkit covering many speech tasks with 200+ training recipes
  • you need to fine-tune HuggingFace pretrained speech models like Whisper, Wav2Vec2, or WavLM
  • you are doing speech research and want customizable, well-documented training pipelines

When to avoid

  • you only need a quick off-the-shelf transcription API with no training or customization
  • your project is not Python/PyTorch based
  • you need production-grade low-latency streaming ASR out of the box without building it yourself

Facets

library · maturity active

speech-recognition audio-processing machine-learning deep-learning tts nlp llm-training speech-processing machine-learning deep-learning artificial-intelligence python cross-platform pytorch conversational-ai asr speaker-recognition speech-enhancement speech-separation huggingface training-recipes diarization natural-language-processing audio gpu

2 sources

Member repositories

RepositoryRoleHealth v2
speechbrain/speechbrainmain83

For agents

markdown · JSON · MCP: product_card(name="speechbrain/speechbrain")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem