clovaai/voxceleb_trainer
In defence of metric learning for speaker recognition observed · 2026-08-28
Health v2 · maintenance only
67/100
- Activity 78
- Release rhythm 35
- Longevity 100
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 2351
- days_rel: n/a
- days_push: 133
- n_releases_24m: 0
Adoption not part of the score
1175 stars · 290 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A PyTorch framework for training and evaluating speaker recognition and verification models on the VoxCeleb datasets. It implements multiple network architectures (ResNet, RawNet3, VGGVox) and metric learning loss functions (AM-Softmax, AAM-Softmax, Angular Prototypical, GE2E, Triplet), with pretrained models provided.
Use cases
- train a speaker verification model on VoxCeleb
- evaluate speaker recognition models with EER metrics
- compare metric learning loss functions for speaker embeddings
- download and run pretrained speaker recognition models
- prepare and augment VoxCeleb audio data for training
- extract speaker embeddings from raw waveforms
When to choose
- you need to train or benchmark speaker verification models on VoxCeleb
- you want to experiment with metric learning losses like AAM-Softmax or Angular Prototypical
- you need reproducible baselines with pretrained models and published EER scores
When to avoid
- you need production-ready speaker recognition as a service rather than a research training framework
- your task is general speech recognition or transcription rather than speaker identity
- you work outside Python/PyTorch environments
Facets
library · maturity maintenance
machine-learning deep-learning audio-processing speech-recognition benchmarking machine-learning deep-learning speech-processing artificial-intelligence python speaker-verification speaker-recognition metric-learning voxceleb pytorch research-code pretrained-models audio linux gpu
1 source
- readme: https://github.com/clovaai/voxceleb_trainer · fetched 2026-08-28 · 2131dfbf70ce
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| clovaai/voxceleb_trainer | main | 67 |
For agents
markdown · JSON · MCP: product_card(name="clovaai/voxceleb_trainer")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem