Ross ROSS = Recommend OSS · open-source software intelligence for agents

TencentGameMate/chinese_speech_pretrain resource

chinese speech pretrained models observed · 2026-08-28

github.com/TencentGameMate/chinese_speech_pretrain · Shell observed · 2026-08-28

Health v2 · maintenance only

32/100

  • Activity 0
  • Release rhythm 35
  • Longevity 100

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1561
  • days_rel: n/a
  • days_push: 740
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1215 stars · 89 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A collection of Chinese speech pretrained models (wav2vec 2.0 and HuBERT, BASE and LARGE) trained by Tencent on 10,000 hours of WenetSpeech data using Fairseq. Checkpoints are hosted on Hugging Face and Baidu Pan, with ESPnet-based ASR recipes demonstrating their use as feature extractors for Conformer speech recognition.

Use cases

  • pretrain wav2vec2 or hubert models on chinese speech
  • use chinese speech representations as features for asr
  • improve mandarin speech recognition with low-resource fine-tuning
  • download chinese wav2vec2 checkpoints for huggingface transformers
  • benchmark chinese speech recognition on aishell and wenetspeech
  • extract self-supervised speech embeddings for downstream audio tasks

When to choose

  • you need chinese-language speech pretrained models for asr or audio feature extraction
  • you want wav2vec2/hubert checkpoints trained on large-scale mandarin data
  • you are fine-tuning speech recognition on limited labeled chinese audio

When to avoid

  • you need pretrained models for non-chinese languages
  • you want a ready-to-use end-to-end speech recognition product rather than pretrained features
  • you cannot work with fairseq or espnet tooling

Facets

dataset · maturity maintenance

speech-recognition machine-learning transformers audio-processing speech-processing machine-learning artificial-intelligence python wav2vec2 hubert self-supervised-pretraining fairseq chinese-asr wenetspeech pretrained-models huggingface natural-language-processing gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
TencentGameMate/chinese_speech_pretrainmain32

For agents

markdown · JSON · MCP: product_card(name="TencentGameMate/chinese_speech_pretrain")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem