Ross ROSS = Recommend OSS · open-source software intelligence for agents

s3prl/s3prl

Self-Supervised Speech Pre-training and Representation Learning Toolkit observed · 2026-08-28

github.com/s3prl/s3prl · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

55/100

  • Activity 71
  • Release rhythm 8
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 2607
  • days_rel: n/a
  • days_push: 174
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

2561 stars · 535 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

S3PRL is a PyTorch toolkit for self-supervised speech pre-training and representation learning, bundling many upstream models like wav2vec 2.0, HuBERT, WavLM, and Mockingjay. It is now in pure maintenance mode, keeping existing functions working while accepting only new upstream model contributions.

Use cases

  • extract speech representations from pretrained self-supervised models
  • fine-tune wav2vec 2.0 or HuBERT for speech recognition
  • benchmark speech models on SUPERB tasks
  • pre-train speech models like APC or TERA
  • evaluate embeddings for speaker verification and emotion recognition
  • run speech downstream tasks like ASR, keyword spotting, and diarization

When to choose

  • you need a unified interface to many self-supervised speech models
  • you want to reproduce SUPERB benchmark results
  • you need pretrained speech encoders for downstream speech tasks

When to avoid

  • you need new features beyond upstream models, since the project is in maintenance mode
  • you want a lightweight production ASR pipeline rather than a research toolkit
  • you need non-speech or text-only self-supervised learning

Facets

library · maturity maintenance

machine-learning speech-recognition audio-processing deep-learning speech-processing machine-learning deep-learning python self-supervised-learning speech-pretraining wav2vec hubert representation-learning pytorch superb audio linux gpu

3 sources

Member repositories

RepositoryRoleHealth v2
s3prl/s3prlmain55

For agents

markdown · JSON · MCP: product_card(name="s3prl/s3prl")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem