modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving. observed · 2026-08-28
Health v2 · maintenance only
99/100
- Activity 99
- Release rhythm 99
- Longevity 98
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 0.0
- age_days: 1379
- days_rel: 7
- days_push: 7
- n_releases_24m: 43
Adoption not part of the score
20036 stars · 2004 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
FunASR is an industrial end-to-end speech recognition toolkit built on PyTorch, offering ASR, VAD, punctuation restoration, speaker diarization, and emotion recognition pipelines for offline, streaming, and edge deployment. It includes OpenAI-compatible serving, WebSocket streaming, vLLM acceleration, and an MCP server for integration with AI agents.
Use cases
- transcribe audio files to text
- real-time streaming speech recognition
- add punctuation to raw transcripts
- identify speakers in meeting recordings
- run speech-to-text on Chinese and multilingual audio
- serve an OpenAI-compatible ASR API
- deploy speech recognition on edge devices
When to choose
- you need production-grade ASR with streaming and VAD pipelines
- you want a Whisper alternative with strong Chinese/multilingual support
- you need speaker diarization and punctuation in one toolkit
- you want OpenAI-compatible or MCP-based speech serving
When to avoid
- you only need text-to-speech synthesis
- you need a lightweight non-PyTorch dependency
- your project requires a non-MIT copyleft-compatible license
Facets
library · maturity active
speech-recognition audio-processing machine-learning llm-inference mcp http-server speech-processing machine-learning python cross-platform cli asr speech-to-text transcription paraformer whisper-alternative voice-activity-detection speaker-diarization punctuation streaming-asr openai-compatible-api vllm gguf websocket chinese multilingual pytorch audio natural-language-processing gpu docker
2 sources
- readme: https://github.com/modelscope/FunASR · fetched 2026-08-28 · 77e6f0ae2270
- registry_pypi: https://pypi.org/pypi/funasr/json · fetched 2026-08-29 · 18ae7779e99c
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| modelscope/FunASR | main | 99 |
For agents
markdown · JSON · MCP: product_card(name="modelscope/FunASR")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem