modelscope/FunASR
在全部 GitHub 仓库 中按Star排名第 141(共 803)
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
项目介绍
Industrial speech recognition toolkit for offline, streaming, and edge deployment.
No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.
Flagship model — Fun-ASR-Nano (LLM-ASR for Chinese, English, and Japanese, plus Chinese dialect groups and regional accents; needs a GPU):
For the separate 31-language checkpoint, use Fun-ASR-MLT-Nano-2512. Language coverage is checkpoint-specific, so Nano and MLT-Nano should be treated as distinct model choices.
On CPU (or for five-language ASR plus emotion and audio-event tags), use SenseVoiceSmall. The pipeline below composes SenseVoiceSmall with FSMN-VAD and CAM++; diarization is provided by…
最新指标
| Star | 20.3k | 2026-09-12 |
|---|---|---|
| Fork | 2.0k | 2026-09-12 |
| 提交 | 5.9k | 2026-09-12 |
| 发布 | 64 | 2026-09-12 |
| Watcher | 118 | 2026-09-12 |
| 开放 issue | 25 | 2026-09-12 |
| 开放 PR | 3 | 2026-09-12 |