GitHubPyPI
FunASR
AI Agent 指数 第 27(共 51) ↓ 下载分享海报
51 综合分
可观测采用度当前有效 · 2026-09-09 49
动量当前有效 · 2026-09-09 55
关注度当前有效 · 2026-09-09 49
信号可信度 依据当前有数据的独立评分维度数量计算。
高3/3 · 2 种数据源
项目介绍
Industrial speech recognition. 170x faster than Whisper. 50+ languages.
No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.
Flagship model — Fun-ASR-Nano (LLM-ASR, 31 languages; the default recommendation, needs a GPU):
On CPU (or for multilingual + emotion in one pass), use SenseVoice — which also returns speaker diarization and timestamps:
Output — structured text with speaker labels, timestamps, and punctuation:
That's it. One model, one call — VAD segmentation, speech recognition, punctuation, speaker diarization all happen automatically.
At scale, accelerate Fun-ASR-Nano with vLLM (batch processing):
Deploy as API…
各数据源
384.9k 月下载量
- 月下载量 384.9k
- 周下载量 57.2k
- 日下载量 9.6k
20.2k Star
- Star 20.2k
- Fork 2.0k
- 提交 5.9k
- 发布 63