TencentGameMate/chinese_speech_pretrain resource
chinese speech pretrained models observed · 2026-08-28
Health v2 · maintenance only
32/100
- Activity 0
- Release rhythm 35
- Longevity 100
Flags: no_releases no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 1561
- days_rel: n/a
- days_push: 740
- n_releases_24m: 0
Adoption not part of the score
1215 stars · 89 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A collection of Chinese speech pretrained models (wav2vec 2.0 and HuBERT, BASE and LARGE) trained by Tencent on 10,000 hours of WenetSpeech data using Fairseq. Checkpoints are hosted on Hugging Face and Baidu Pan, with ESPnet-based ASR recipes demonstrating their use as feature extractors for Conformer speech recognition.
Use cases
- pretrain wav2vec2 or hubert models on chinese speech
- use chinese speech representations as features for asr
- improve mandarin speech recognition with low-resource fine-tuning
- download chinese wav2vec2 checkpoints for huggingface transformers
- benchmark chinese speech recognition on aishell and wenetspeech
- extract self-supervised speech embeddings for downstream audio tasks
When to choose
- you need chinese-language speech pretrained models for asr or audio feature extraction
- you want wav2vec2/hubert checkpoints trained on large-scale mandarin data
- you are fine-tuning speech recognition on limited labeled chinese audio
When to avoid
- you need pretrained models for non-chinese languages
- you want a ready-to-use end-to-end speech recognition product rather than pretrained features
- you cannot work with fairseq or espnet tooling
Facets
dataset · maturity maintenance
speech-recognition machine-learning transformers audio-processing speech-processing machine-learning artificial-intelligence python wav2vec2 hubert self-supervised-pretraining fairseq chinese-asr wenetspeech pretrained-models huggingface natural-language-processing gpu linux
1 source
- readme: https://github.com/TencentGameMate/chinese_speech_pretrain · fetched 2026-08-28 · 18750fd41ff3
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| TencentGameMate/chinese_speech_pretrain | main | 32 |
For agents
markdown · JSON · MCP: product_card(name="TencentGameMate/chinese_speech_pretrain")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem