aiagent.club
English
GitHub 仓库

modelscope/FunASR

Python 首次收录 2022-11-24 更新于 2026-09-12 查看源 ↗

在全部 GitHub 仓库 中按Star排名第 141(共 803)

20.3k Star

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

项目介绍

Industrial speech recognition toolkit for offline, streaming, and edge deployment.

No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.

Flagship model — Fun-ASR-Nano (LLM-ASR for Chinese, English, and Japanese, plus Chinese dialect groups and regional accents; needs a GPU):

For the separate 31-language checkpoint, use Fun-ASR-MLT-Nano-2512. Language coverage is checkpoint-specific, so Nano and MLT-Nano should be treated as distinct model choices.

On CPU (or for five-language ASR plus emotion and audio-event tags), use SenseVoiceSmall. The pipeline below composes SenseVoiceSmall with FSMN-VAD and CAM++; diarization is provided by…

摘自 github.com/modelscope/FunASR

最新指标

Star 20.3k 2026-09-12
Fork 2.0k 2026-09-12
提交 5.9k 2026-09-12
发布 64 2026-09-12
Watcher 118 2026-09-12
开放 issue 25 2026-09-12
开放 PR 3 2026-09-12