GitHubPyPI
FunASR
AI Agent 指数 第 28(共 51) ↓ 下载分享海报
51 综合分
可观测采用度当前有效 · 2026-09-12 51
动量当前有效 · 2026-09-12 53
关注度当前有效 · 2026-09-12 49
信号可信度 依据当前有数据的独立评分维度数量计算。
高3/3 · 2 种数据源
项目介绍
Industrial speech recognition. 170x faster than Whisper. 50+ languages.
No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.
Flagship model — Fun-ASR-Nano (LLM-ASR, 31 languages; the default recommendation, needs a GPU):
On CPU (or for multilingual + emotion in one pass), use SenseVoice — which also returns speaker diarization and timestamps:
Output — structured text with speaker labels, timestamps, and punctuation:
That's it. One model, one call — VAD segmentation, speech recognition, punctuation, speaker diarization all happen automatically.
At scale, accelerate Fun-ASR-Nano with vLLM (batch processing):
Deploy as API…
各数据源
355.6k 月下载量
- 月下载量 355.6k
- 周下载量 58.7k
- 日下载量 9.5k
20.3k Star
- Star 20.3k
- Fork 2.0k
- 提交 5.9k
- 发布 64